Skip to content
Advertisement

Dataset Card for DiaBLa: Bilingual dialogue parallel evaluation set Dataset Summary The dataset is an English-French dataset for the evaluation of Machine Translation (MT) for informal, written bilingual dialogue. The dataset contains 144 spontaneous dialogues (5,700+ sentences) between native English and French speakers, mediated by one of two neural MT systems in a range of role-play settings. See below for some basic statistics. The dialogues are accompanied by… See the full description on the dataset page:

Source: Hugging Face Hub (rbawden/DiaBLa). Metadata imported from the dataset’s Hub tags.

Advertisement