Skip to content
Advertisement
Text

United Nations Parallel Corpus

United Nations Parallel Corpus

Dataset Card for United Nations Parallel Corpus Dataset Summary The United Nations Parallel Corpus is the first parallel corpus composed from United Nations documents published by the original data creator. The parallel corpus consists of manually translated UN documents from the last 25 years (1990 to 2014) for the six official UN languages, Arabic, Chinese, English, French, Russian, and Spanish. The corpus is freely available for download under a liberal license.… See the full description on the dataset page: pc.

Source: Hugging Face Hub (Helsinki-NLP/un_pc). Metadata imported from the dataset’s Hub tags.

Advertisement