Skip to content
Advertisement

IMPaCTS: Italian Multi-level Parallel Corpus for Controlled Text Simplification IMPaCTS is a large-scale Italian parallel corpus for controlled text simplification, containing complex–simple sentence pairs automatically generated using Large Language Models. Each pair is annotated with readability scores (via Read-IT; paper here) and a rich set of linguistic features obtained with ProfilingUD (paper here, web-based tool here). The dataset is a cleaned subset of the dataset… See the full description on the dataset page:

Source: Hugging Face Hub (mpapucci/impacts). Metadata imported from the dataset’s Hub tags.

Advertisement