Skip to content
Advertisement

10M–100M

169 results

Text

WMT14

WMT14

10M–100M·Custom / Research-only·Parquet
ImageMultimodalText

S1-MMAlign

S1-MMAlign A Large-Scale Multi-Disciplinary Scientific Multimodal Dataset S1-MMAlign is a large-scale, multi-disciplinary multimodal dataset…

10M–100M·CC-BY-NC·WebDataset
ImageMultimodalText

MAmmoTH-VL-Instruct-12M

MAmmoTH-VL-Instruct-12M 🏠 Homepage 🤖 MAmmoTH-VL-8B 💻 Code 📄 Arxiv 📕 PDF 🖥️ Demo Introduction Our simple yet scalable…

10M–100M·Apache-2.0·WebDataset
MultimodalTabularText

UniST

UniST This dataset contains UniST codec-token training data exported from local metadata and codec results. We train UniSS…

10M–100M·CC-BY-NC
Audio

LibriQuote

This repository contains the LibriQuote dataset, a speech dataset of fictional character utterances for expressive zero-shot speech synthesis.…

10M–100M·CC-BY-NC
Text

CapSpeech

CapSpeech DataSet used for the paper: CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech Please refer to CapSpeech repo…

10M–100M·CC-BY-NC·Parquet
AudioMultimodalText

Pretraining-V1

Indic TTS Unified v1 A large-scale, unified collection of speech data for text-to-speech (TTS) and speech research. This…

10M–100M·CC-BY·Parquet
MultimodalSensor / Time-seriesTabular

droid

This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v3.0", "robot type": "franka", "total episodes":…

10M–100M·Apache-2.0·Parquet
Advertisement