datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
allmusiccaps
AllMusicCaps
Music caption dataset built from professional AllMusic album reviews, cross-referenced with Discogs
releases and YouTube tracks. Captions are generated in two complementary styles by a two-stage LLM
pipeline.
Released with the ISMIR 2026 paper AllMusicCaps: Album Reviews as Complementary Supervision for
Music CLAP. Code and models: github.com/MTG/allmusiccaps.
This dataset contains no audio, only identifiers and captions. Recover audio from the
youtube_id field, or… See the full description on the dataset page: https://huggingface.co/datasets/mtg-upf/allmusiccaps.mtg-vlm-training-dataAgri-VLM-VectorsMTGroupDatasetMTG_Rulebookfine_tune_mtg_otjThis is a TESTING data set of question/answer pairs about a recent MTG card set to practice and test fine tuning. It has not been thoroughly vetted and probably has bad data.
zelk12__MT-Merge2-MU-gemma-2-MTg2MT1g2-9B-details
Dataset Card for Evaluation run of zelk12/MT-Merge2-MU-gemma-2-MTg2MT1g2-9B
Dataset automatically created during the evaluation run of model zelk12/MT-Merge2-MU-gemma-2-MTg2MT1g2-9B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/zelk12__MT-Merge2-MU-gemma-2-MTg2MT1g2-9B-details.MTG_cards
