Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
CORAA-v1.1 CORAA-v1.1 is a publicly available dataset for Automatic Speech Recognition (ASR) in the Brazilian Portuguese language containing 290.77 hours of audios and their respective transcriptions (400k+ segmented audios). The dataset is composed of audios of 5 original projects: ALIP (Gonçalves, 2019) C-ORAL Brazil (Raso and Mello, 2012) NURC-Recife (Oliviera Jr., 2016) SP-2010 (Mendes and Oushiro, 2012) TEDx talks (talks in Portuguese) The audios were either validated by… See the full description on the dataset page:
Source: Hugging Face Hub (Racoci/CORAA-v1.1). Metadata imported from the dataset’s Hub tags.