Skip to content
Advertisement
Text

The Tatoeba Translation Challenge

The Tatoeba Translation Challenge

Dataset Card for DigitalLearningGmbH/tatoeba mt parquet This is a mirror of Helsinki-NLP/tatoeba mt, converted to parquet for compatibility with newer huggingface requirements. Original dataset card follows. Dataset Summary The Tatoeba Translation Challenge is a multilingual data set of machine translation benchmarks derived from user-contributed translations collected by Tatoeba.org and provided as parallel corpus from OPUS. This dataset includes test and development… See the full description on the dataset page: mt parquet.

Source: Hugging Face Hub (DigitalLearningGmbH/tatoeba_mt_parquet). Metadata imported from the dataset’s Hub tags.

Advertisement