Filters
Search
Product type
Language
Country
Year of Collection

Arabic NER news text

More info
Common Use CasesNER, Content Classification, Search Engines
Dataset IDARB_NER001
TypeText
Unit20,774 sentences
LanguageArabic (Standard)
CountryN/A

Dari (Afghanistan) broadcast

More info
Common Use CasesASR, Automatic Captioning, Keyword Spotting
Dataset IDDAR_BRC001
TypeAudio
Unit49 hours
LanguageDari
CountryAfghanistan

East African facial images

More info
Common Use CasesFacial Recognition
Dataset IDIMG_FACE_KEN_CN
TypeImage
Unit13500 images
LanguageN/A
CountryKenya

English Inverse text normalisation

More info
Common Use CasesASR, Language Modelling, Closed Captioning
Dataset IDENG_ITN001
TypeText
Unit4454 test cases
LanguageEnglish
CountryN/A

English NER news text

More info
Common Use CasesNER, Content Classification, Search Engines
Dataset IDENG_NER001
TypeText
Unit22,768 sentences
LanguageEnglish
CountryN/A

European License Plate Detection Annotations

More info
Common Use CasesLicense plate detection for vehicles on the road
Dataset IDLICENSE_ANNO
TypeImage Annotation
Unit100,000 bounding boxes
LanguageN/A
CountryGermany, France, Switzerland

Get Started with Off-the-Shelf AI Training Datasets

Appen’s extensive catalog of off-the-shelf (OTS) datasets spans multiple data types and industries, providing comprehensive coverage for various AI applications. These datasets are crafted to the highest standards of quality and accuracy, ensuring reliable training data for AI models.

Talk to an expert