Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Summary Vox Classica is a Latin speech corpus of ~73 hours of audio, segmented into short audio clips by sentence. Vox Classica is a large-scale, ML-ready dataset of human-read Classical Latin. It was designed to address the absence of a publicly available human-read Latin corpus large enough for model training. Alignment and curation: Kaiyuan Zhao Language: Latin (Classical) Uses This dataset is built for training and evaluating speech processing models for… See the full description on the dataset page:
Source: Hugging Face Hub (Ken-Z/Latin-Audio). Metadata imported from the dataset’s Hub tags.