datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
transformers-merge-experimentsMoS-Experiment-Data-Archive
MoS Experiment Data Archive
Public data archive for the DFlash / Aurora MoS experiments.
Contents:
dom250k/ and dom250k_train/: domain-specialist training data.
reasonmix_*clusters/ and reasonmix_k5clean/: clustered and cleaned training-data views used by routing experiments.
natclusters/: natural-cluster data view.
gen800k/: current 800K large-data experiment inputs. This copy remains on Weka until the active 800K experiment is complete.
Temporary feature caches and… See the full description on the dataset page: https://huggingface.co/datasets/ryan-0608/MoS-Experiment-Data-Archive.lm-eval-results-automerger-Experiment28Yam-7B-private
Dataset Card for Evaluation run of automerger/Experiment28Yam-7B
Dataset automatically created during the evaluation run of model automerger/Experiment28Yam-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-automerger-Experiment28Yam-7B-private.zoya-image-1-experiments
ZOYA IMAGE-1 — Reproducible GGUF Experiments
Purpose
This dataset stores reproducible ZOYA IMAGE-1 image-generation
experiments together with the exact generation parameters,
model identities, SHA256 fingerprints, and validation reports.
The package is designed for controlled comparisons where the
tested variable is changed explicitly and all other relevant
variables remain fixed.
Current baseline
Experiment ID: ZOYA_PHASE0_BASELINE_00001… See the full description on the dataset page: https://huggingface.co/datasets/tigerking009/zoya-image-1-experiments.ml-experiment-corpuslm-eval-results-MaziyarPanahi-YamshadowInex12_Experiment26T3q-private
Dataset Card for Evaluation run of MaziyarPanahi/YamshadowInex12_Experiment26T3q
Dataset automatically created during the evaluation run of model MaziyarPanahi/YamshadowInex12_Experiment26T3q
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-YamshadowInex12_Experiment26T3q-private.lm-eval-results-MaziyarPanahi-MeliodasPercival_01_Experiment26T3q-private
Dataset Card for Evaluation run of MaziyarPanahi/MeliodasPercival_01_Experiment26T3q
Dataset automatically created during the evaluation run of model MaziyarPanahi/MeliodasPercival_01_Experiment26T3q
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-MeliodasPercival_01_Experiment26T3q-private.abmelt-experiments-exp_20260220_130124adaption-experimentalcotqa
This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform.
adaption-ExperimentalCoTQA
This dataset contains 1,000 question-and-answer pairs focusing on the evolution of African societies, including topics like the role of women, indigenous languages, traditional leadership, and polyrhythmic music. The content explores historical transitions from pre-colonial eras through colonialism to modern-day challenges and cultural innovations.… See the full description on the dataset page: https://huggingface.co/datasets/Svngoku/adaption-experimentalcotqa.tr-cc-experimental
Veri Seti Hakkında
Kaynak: Common Crawl (CC) Türkçe Ağ Verileri
Toplayan [Ben (Swag Victoria)]
Temizleme & Filtreleme: GLM 5.3 Flash yardımıyla HTML etiketleri, navigasyon gürültüleri ve spam metinler ayıklanmaya çalışılmıştır.
Format: .jsonl (JSON Lines)
Q&A Abi neden bu kadar yavaş geliyor CC kazıma işlemi? (Answer) Abi'nin amına koyim.
lm-eval-results-MaziyarPanahi-Experiment26Yamshadow_Ognoexperiment27Multi_verse_model-private
Dataset Card for Evaluation run of MaziyarPanahi/Experiment26Yamshadow_Ognoexperiment27Multi_verse_model
Dataset automatically created during the evaluation run of model MaziyarPanahi/Experiment26Yamshadow_Ognoexperiment27Multi_verse_model
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-Experiment26Yamshadow_Ognoexperiment27Multi_verse_model-private.fm-model-experiments-data
FM Model Experiments — Synthetic VLM Training Data (KO + EN)
Annotation data produced while building a native-resolution Korean+English VLM
(GLM-4.6V vision tower transplanted onto a frozen GLM-5.2 743B MoE decoder).
Code + technical report: https://github.com/genonai/fm-model-experiments
This repo contains ANNOTATIONS ONLY (.jsonl). No images are redistributed.
Every row references an image by a relative path (data/...); obtain the images
from the original sources listed below… See the full description on the dataset page: https://huggingface.co/datasets/mncai/fm-model-experiments-data.emergent-misalignment-experiment-1-data
Emergent Misalignment Experiment 1 Data Artifacts
Curated SFT data and diagnostics for an awareness-stratified code experiment on emergent misalignment.
This artifact contains the exact trainable JSONL branches used for the reported n=1000 and n=3452 runs, plus the small manifests and balance summaries needed to audit the data mixture. The paired model adapters are available at jash404/emergent-misalignment-experiment-1-adapters. The source code and reports are in… See the full description on the dataset page: https://huggingface.co/datasets/jash404/emergent-misalignment-experiment-1-data.lm-eval-results-MaziyarPanahi-Experiment26Yam_Ognoexperiment27Multi_verse_model-private
Dataset Card for Evaluation run of MaziyarPanahi/Experiment26Yam_Ognoexperiment27Multi_verse_model
Dataset automatically created during the evaluation run of model MaziyarPanahi/Experiment26Yam_Ognoexperiment27Multi_verse_model
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-Experiment26Yam_Ognoexperiment27Multi_verse_model-private.abmelt-experiments-exp_20260220_125440lm-eval-results-MaziyarPanahi-M7Yamshadowexperiment28_Experiment26T3q-private
Dataset Card for Evaluation run of MaziyarPanahi/M7Yamshadowexperiment28_Experiment26T3q
Dataset automatically created during the evaluation run of model MaziyarPanahi/M7Yamshadowexperiment28_Experiment26T3q
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-M7Yamshadowexperiment28_Experiment26T3q-private.lm-eval-results-MaziyarPanahi-YamshadowStrangemerges_32_Experiment24Ognoexperiment27-private
Dataset Card for Evaluation run of MaziyarPanahi/YamshadowStrangemerges_32_Experiment24Ognoexperiment27
Dataset automatically created during the evaluation run of model MaziyarPanahi/YamshadowStrangemerges_32_Experiment24Ognoexperiment27
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-YamshadowStrangemerges_32_Experiment24Ognoexperiment27-private.lm-eval-results-automerger-Experiment29Pastiche-7B-private
Dataset Card for Evaluation run of automerger/Experiment29Pastiche-7B
Dataset automatically created during the evaluation run of model automerger/Experiment29Pastiche-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-automerger-Experiment29Pastiche-7B-private.dataset_fenics_experiment_v7grok-1-ternary-quant-experiments
Grok-1 SAAQ quantization / route-preservation experiments
Dataset author: Raul Montoya Cardenas (rmems)
SAAQ stands for Spiking Adaptive Activity Quantization, a term coined by
the dataset author.
Attribution: Grok Build: Grok 4.5 (high) packaged the original 2026-08-10
dataset. Codex: GPT-5.6-Sol (OpenAI) implemented, executed, validated, and published the
canonical issue #85 v4 evidence added on 2026-08-24.
Personal research measuring route preservation when packing open… See the full description on the dataset page: https://huggingface.co/datasets/rmems/grok-1-ternary-quant-experiments.abmelt-experiments-exp_20260218_171120experiment-001-budget-boxed
KamiBench Experiment 001 — budget-boxed agents in Kamigotchi
Complete agentic traces from experiment 001, the KamiBench
calibration run: three LLM agents dropped into
Kamigotchi, a live, persistent, on-chain
world (Yominet), each with a $10 inference budget, a 7-day
wall-clock cap, and no further human contact. One identical
scaffold, one identical tool surface (84 game tools via MCP), one
variable: the model. The agents schedule their own wake-ups, keep
their own files, and act… See the full description on the dataset page: https://huggingface.co/datasets/KamiBench/experiment-001-budget-boxed.rl-experiment-rescue-lfsmedical_dialogue_swe_experiment1Synthetic dataset of medical encounters in Swedish generated with Llama 3 70B.
The conversations do not represent ideal communication but strive to capture conversations as they may happen in medical settings. Patients have a randomly sampled personality with assumed decorrelated normal distribution based on the psychometric data from Twenty-item, Table 3 in Ringwald, W. R. et al (2022). Psychometric evaluation of a Big Five personality state scale for intensive longitudinal studies.… See the full description on the dataset page: https://huggingface.co/datasets/Debbevi/medical_dialogue_swe_experiment1.abmelt-experiments-exp_20260219_182250persona-belief-probesexperimental-optimizationabmelt-experiments-exp_20260216_181056lm-eval-results-ChaoticNeutrals-Prima-LelantaclesV7-experimental-7b-private
Dataset Card for Evaluation run of ChaoticNeutrals/Prima-LelantaclesV7-experimental-7b
Dataset automatically created during the evaluation run of model ChaoticNeutrals/Prima-LelantaclesV7-experimental-7b
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-ChaoticNeutrals-Prima-LelantaclesV7-experimental-7b-private.abmelt-experiments-exp_20260219_182408
