Skip to content
Advertisement

Dataset Card for : Arabic Aya (2A) Arabic Aya (2A) : A Curated Subset of the Aya Collection for Arabic Language Processing Dataset Sources & Infos Data Origin: Derived from 69 subsets of the original Aya datasets : CohereForAI/aya collection, CohereForAI/aya dataset, and CohereForAI/aya evaluation suite. Languages: Modern Standard Arabic (MSA) and a variety of Arabic dialects ( ‘arb’, ‘arz’, ‘ary’, ‘ars’, ‘knc’, ‘acm’, ‘apc’, ‘aeb’, ‘ajp’… See the full description on the dataset page: Aya.

Source: Hugging Face Hub (yrrhall/Arabic_Aya). Metadata imported from the dataset’s Hub tags.

Advertisement