Apache-2.0
807 results
Random Erasing Data Augmentation. Experiments on CIFAR10, CIFAR100 and Fashion-MNIST
Image
Open Source
A Implementation of SpecAugment with Tensorflow & Pytorch, introduced by Google Brain
Audio
Open Source
A complete end-to-end demonstration in which we collect training data in Unity and use that data to train a deep…
Image
Open Source
Generate High-Quality Synthetics, Train, Measure, and Evaluate in a Single Pipeline
Multimodal
Open Source
Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipeline
Text
Open Source
SDG is a specialized framework designed to generate high-quality structured tabular data.
Text
Open Source
🎨 NeMo Data Designer: Generate high-quality synthetic data from scratch or from seed data.
Text
Open Source