datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Speech-IFEvalAudioX-IFcaps
[ICLR 2026] AudioX-IFcaps: Instruction-Following Audio Caption Dataset
AudioX-IFcaps (Instruction-Following) is a large-scale, high-quality multimodal dataset designed for training unified audio and music generation models. The dataset contains over 7 million samples with fine-grained, structured annotations that enable precise control over audio generation, including sound event categories, counts, temporal ordering, and timestamps.
📊 Dataset Statistics
General Audio:… See the full description on the dataset page: https://huggingface.co/datasets/HKUSTAudio/AudioX-IFcaps.IfGPT-OPEN-Dataset
IfGPT Dataset
Objectives of the project IfGPT
The IfGPT Dataset is developed within the project IfGPT: Infrastructure for Fine-tuning Pre-trained Large Language Models which aims to establish a freely accessible infrastructure for the selection and pre-processing of large datasets for Bulgarian as well as tailored data for particular industries and fine-tuning suitable freely available large language models for specific purposes.
IfGPT Dataset
IfGPT… See the full description on the dataset page: https://huggingface.co/datasets/DCL-IBL/IfGPT-OPEN-Dataset.
