Skip to content
Advertisement
Text

NusaTranslationBitextMining

NusaTranslationBitextMining An MTEB dataset Massive Text Embedding Benchmark NusaTranslation is a parallel dataset for machine translation on 11…

NusaTranslationBitextMining An MTEB dataset Massive Text Embedding Benchmark NusaTranslation is a parallel dataset for machine translation on 11 Indonesia languages and English. Task category t2t Domains Social, Written Reference mt How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task =… See the full description on the dataset page:

Source: Hugging Face Hub (mteb/NusaTranslationBitextMining). Metadata imported from the dataset’s Hub tags.

Advertisement