Skip to content
Advertisement
Text

chr_en

Dataset Card for ChrEn Dataset Summary ChrEn is a Cherokee-English parallel dataset to facilitate machine translation research between Cherokee and…

Dataset Card for ChrEn Dataset Summary ChrEn is a Cherokee-English parallel dataset to facilitate machine translation research between Cherokee and English. ChrEn is extremely low-resource contains 14k sentence pairs in total, split in ways that facilitate both in-domain and out-of-domain evaluation. ChrEn also contains 5k Cherokee monolingual data to enable semi-supervised learning. Supported Tasks and Leaderboards The dataset is intended to use for… See the full description on the dataset page: en.

Source: Hugging Face Hub (shiyue/chr_en). Metadata imported from the dataset’s Hub tags.

Advertisement