Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Crowd Whatsapp Yiddish - Source Dataset
About This dataset was created by crowd-sourced Whatsapp voice recordings in Yiddish as part of the ivrit.ai project. Volunteers read a message sent to them from a predefined set of messages, recording themselves using Whasapp voice message sent to the collecting bot. Later this data is normalized by aligning the captions with the audio using Stable Whisper (See Below). The recording project is an ongoing effort and new data will be appended to this dataset periodically as it is… See the full description on the dataset page:
Source: Hugging Face Hub (ivrit-ai/crowd-whatsapp-yi). Metadata imported from the dataset’s Hub tags.