datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cocktail-party-mixtures
Cocktail-party mixtures with an arbitrary number of sources
Frozen evaluation sets from
DJRHails/voxpath (voxpath.cocktail,
docs/cocktail-party-separation-2026-09.md). Each mixture has one to four
talkers, optionally one music track and one noise track, placed in a simulated
room (pyroomacoustics image-source method, RT60 0.2 to 0.8 s) and recorded by one
microphone. Every file is mono 16 kHz.
set
mixtures
seconds
talkers
music / noise
reverberant
val
300
4
1 to 4
147… See the full description on the dataset page: https://huggingface.co/datasets/DJRHails/cocktail-party-mixtures.math_rlvr_mixture_dpomath_rlvr_mixture_sftjsfs11__MixtureofMerges-MoE-4x7b-v4-details
Dataset Card for Evaluation run of jsfs11/MixtureofMerges-MoE-4x7b-v4
Dataset automatically created during the evaluation run of model jsfs11/MixtureofMerges-MoE-4x7b-v4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/jsfs11__MixtureofMerges-MoE-4x7b-v4-details.jsfs11__MixtureofMerges-MoE-4x7b-v5-details
Dataset Card for Evaluation run of jsfs11/MixtureofMerges-MoE-4x7b-v5
Dataset automatically created during the evaluation run of model jsfs11/MixtureofMerges-MoE-4x7b-v5
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/jsfs11__MixtureofMerges-MoE-4x7b-v5-details.mixture-of-experts-papers
Mixture of Experts Papers — FineSet
A research-paper dataset on Mixture of Experts Papers, assembled, deduplicated, and quality-scored by
FineSet from arXiv and Semantic Scholar.
📸 This is a dated snapshot — generated 2026-06-19.
It is not auto-updated. Research on Mixture of Experts Papers moves fast — new papers land on arXiv every
week. Want this same dataset refreshed daily, on a topic you choose? See the bottom. ↓
Why this dataset
Quality-scored:… See the full description on the dataset page: https://huggingface.co/datasets/fineset-io/mixture-of-experts-papers.code_rlvr_mixture_dpocode_rlvr_mixture_sftmixture-then-select-selections
Frozen Qwen3-8B Selections
This dataset contains the exact selection metadata and pool indices used for
a frozen cross-scale data-selection experiment. The subsets were selected
using Qwen3-8B-derived information and can be transferred unchanged to a
tokenizer-compatible larger target model.
The companion training and evaluation code is:
https://github.com/submissionpaper1234/mixture-then-select-reproducibility
Critical Interpretation
The selected instruction text… See the full description on the dataset page: https://huggingface.co/datasets/submissionpaper1234/mixture-then-select-selections.
