Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Card for Tamil Speech Dataset Summary This dataset consists of 7 hours of transcribed high-quality audio of Chilean Spanish sentences recorded by 31 volunteers. The dataset is intended for speech technologies. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. Supported Tasks text-to-speech, text-to-audio: The dataset can be used to train a model for Text-To-Speech (TTS). automatic-speech-recognition… See the full description on the dataset page:
Source: Hugging Face Hub (ylacombe/google-chilean-spanish). Metadata imported from the dataset’s Hub tags.