CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01evalstate /transformers-merge-experimentstabularn<1K3 likes1.8k downloads5mo agoHugging Face02ryan-0608 /MoS-Experiment-Data-Archive MoS Experiment Data Archive Public data archive for the DFlash / Aurora MoS experiments. Contents: dom250k/ and dom250k_train/: domain-specialist training data. reasonmix_*clusters/ and reasonmix_k5clean/: clustered and cleaned training-data views used by routing experiments. natclusters/: natural-cluster data view. gen800k/: current 800K large-data experiment inputs. This copy remains on Weka until the active 800K experiment is complete. Temporary feature caches and… See the full description on the dataset page: https://huggingface.co/datasets/ryan-0608/MoS-Experiment-Data-Archive.text100K<n<1M0 likes435 downloads2mo agoHugging Face03nyu-dice-lab /lm-eval-results-automerger-Experiment28Yam-7B-private Dataset Card for Evaluation run of automerger/Experiment28Yam-7B Dataset automatically created during the evaluation run of model automerger/Experiment28Yam-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-automerger-Experiment28Yam-7B-private.tabular100K<n<1M0 likes257 downloads2y agoHugging Face04tigerking009 /zoya-image-1-experiments ZOYA IMAGE-1 — Reproducible GGUF Experiments Purpose This dataset stores reproducible ZOYA IMAGE-1 image-generation experiments together with the exact generation parameters, model identities, SHA256 fingerprints, and validation reports. The package is designed for controlled comparisons where the tested variable is changed explicitly and all other relevant variables remain fixed. Current baseline Experiment ID: ZOYA_PHASE0_BASELINE_00001… See the full description on the dataset page: https://huggingface.co/datasets/tigerking009/zoya-image-1-experiments.imageimage-to-imagen<1K0 likes133 downloads1mo agoHugging Face05ArUn123123 /ml-experiment-corpustextn<1K0 likes125 downloads29d agoHugging Face06nyu-dice-lab /lm-eval-results-MaziyarPanahi-YamshadowInex12_Experiment26T3q-private Dataset Card for Evaluation run of MaziyarPanahi/YamshadowInex12_Experiment26T3q Dataset automatically created during the evaluation run of model MaziyarPanahi/YamshadowInex12_Experiment26T3q The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-YamshadowInex12_Experiment26T3q-private.tabular100K<n<1M0 likes118 downloads2y agoHugging Face07nyu-dice-lab /lm-eval-results-MaziyarPanahi-MeliodasPercival_01_Experiment26T3q-private Dataset Card for Evaluation run of MaziyarPanahi/MeliodasPercival_01_Experiment26T3q Dataset automatically created during the evaluation run of model MaziyarPanahi/MeliodasPercival_01_Experiment26T3q The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-MeliodasPercival_01_Experiment26T3q-private.tabular100K<n<1M0 likes108 downloads2y agoHugging Face08Praful932 /abmelt-experiments-exp_20260220_130124textn<1K0 likes93 downloads7mo agoHugging Face09Svngoku /adaption-experimentalcotqa This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform. adaption-ExperimentalCoTQA This dataset contains 1,000 question-and-answer pairs focusing on the evolution of African societies, including topics like the role of women, indigenous languages, traditional leadership, and polyrhythmic music. The content explores historical transitions from pre-colonial eras through colonialism to modern-day challenges and cultural innovations.… See the full description on the dataset page: https://huggingface.co/datasets/Svngoku/adaption-experimentalcotqa.text1K<n<10K1 likes90 downloads1mo agoHugging Face10Swagvictoria /tr-cc-experimental Veri Seti Hakkında Kaynak: Common Crawl (CC) Türkçe Ağ Verileri Toplayan [Ben (Swag Victoria)] Temizleme & Filtreleme: GLM 5.3 Flash yardımıyla HTML etiketleri, navigasyon gürültüleri ve spam metinler ayıklanmaya çalışılmıştır. Format: .jsonl (JSON Lines) Q&A Abi neden bu kadar yavaş geliyor CC kazıma işlemi? (Answer) Abi'nin amına koyim. textn<1K0 likes89 downloads10d agoHugging Face11nyu-dice-lab /lm-eval-results-MaziyarPanahi-Experiment26Yamshadow_Ognoexperiment27Multi_verse_model-private Dataset Card for Evaluation run of MaziyarPanahi/Experiment26Yamshadow_Ognoexperiment27Multi_verse_model Dataset automatically created during the evaluation run of model MaziyarPanahi/Experiment26Yamshadow_Ognoexperiment27Multi_verse_model The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-Experiment26Yamshadow_Ognoexperiment27Multi_verse_model-private.tabular100K<n<1M0 likes78 downloads2y agoHugging Face12mncai /fm-model-experiments-data FM Model Experiments — Synthetic VLM Training Data (KO + EN) Annotation data produced while building a native-resolution Korean+English VLM (GLM-4.6V vision tower transplanted onto a frozen GLM-5.2 743B MoE decoder). Code + technical report: https://github.com/genonai/fm-model-experiments This repo contains ANNOTATIONS ONLY (.jsonl). No images are redistributed. Every row references an image by a relative path (data/...); obtain the images from the original sources listed below… See the full description on the dataset page: https://huggingface.co/datasets/mncai/fm-model-experiments-data.textvisual-question-answering1K<n<10K0 likes77 downloads3mo agoHugging Face13jash404 /emergent-misalignment-experiment-1-data Emergent Misalignment Experiment 1 Data Artifacts Curated SFT data and diagnostics for an awareness-stratified code experiment on emergent misalignment. This artifact contains the exact trainable JSONL branches used for the reported n=1000 and n=3452 runs, plus the small manifests and balance summaries needed to audit the data mixture. The paired model adapters are available at jash404/emergent-misalignment-experiment-1-adapters. The source code and reports are in… See the full description on the dataset page: https://huggingface.co/datasets/jash404/emergent-misalignment-experiment-1-data.tabulartext-generationn<1K0 likes75 downloads4mo agoHugging Face14nyu-dice-lab /lm-eval-results-MaziyarPanahi-Experiment26Yam_Ognoexperiment27Multi_verse_model-private Dataset Card for Evaluation run of MaziyarPanahi/Experiment26Yam_Ognoexperiment27Multi_verse_model Dataset automatically created during the evaluation run of model MaziyarPanahi/Experiment26Yam_Ognoexperiment27Multi_verse_model The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-Experiment26Yam_Ognoexperiment27Multi_verse_model-private.tabular100K<n<1M0 likes69 downloads2y agoHugging Face15Praful932 /abmelt-experiments-exp_20260220_125440textn<1K0 likes69 downloads7mo agoHugging Face16nyu-dice-lab /lm-eval-results-MaziyarPanahi-M7Yamshadowexperiment28_Experiment26T3q-private Dataset Card for Evaluation run of MaziyarPanahi/M7Yamshadowexperiment28_Experiment26T3q Dataset automatically created during the evaluation run of model MaziyarPanahi/M7Yamshadowexperiment28_Experiment26T3q The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-M7Yamshadowexperiment28_Experiment26T3q-private.tabular100K<n<1M0 likes68 downloads2y agoHugging Face17nyu-dice-lab /lm-eval-results-MaziyarPanahi-YamshadowStrangemerges_32_Experiment24Ognoexperiment27-private Dataset Card for Evaluation run of MaziyarPanahi/YamshadowStrangemerges_32_Experiment24Ognoexperiment27 Dataset automatically created during the evaluation run of model MaziyarPanahi/YamshadowStrangemerges_32_Experiment24Ognoexperiment27 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MaziyarPanahi-YamshadowStrangemerges_32_Experiment24Ognoexperiment27-private.tabular100K<n<1M0 likes68 downloads2y agoHugging Face18nyu-dice-lab /lm-eval-results-automerger-Experiment29Pastiche-7B-private Dataset Card for Evaluation run of automerger/Experiment29Pastiche-7B Dataset automatically created during the evaluation run of model automerger/Experiment29Pastiche-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-automerger-Experiment29Pastiche-7B-private.tabular100K<n<1M0 likes66 downloads2y agoHugging Face19orange67 /dataset_fenics_experiment_v7text1K<n<10K2 likes60 downloads1y agoHugging Face20rmems /grok-1-ternary-quant-experiments Grok-1 SAAQ quantization / route-preservation experiments Dataset author: Raul Montoya Cardenas (rmems) SAAQ stands for Spiking Adaptive Activity Quantization, a term coined by the dataset author. Attribution: Grok Build: Grok 4.5 (high) packaged the original 2026-08-10 dataset. Codex: GPT-5.6-Sol (OpenAI) implemented, executed, validated, and published the canonical issue #85 v4 evidence added on 2026-08-24. Personal research measuring route preservation when packing open… See the full description on the dataset page: https://huggingface.co/datasets/rmems/grok-1-ternary-quant-experiments.tabularothern<1K0 likes60 downloads1mo agoHugging Face21Praful932 /abmelt-experiments-exp_20260218_171120textn<1K0 likes59 downloads7mo agoHugging Face22KamiBench /experiment-001-budget-boxed KamiBench Experiment 001 — budget-boxed agents in Kamigotchi Complete agentic traces from experiment 001, the KamiBench calibration run: three LLM agents dropped into Kamigotchi, a live, persistent, on-chain world (Yominet), each with a $10 inference budget, a 7-day wall-clock cap, and no further human contact. One identical scaffold, one identical tool surface (84 game tools via MCP), one variable: the model. The agents schedule their own wake-ups, keep their own files, and act… See the full description on the dataset page: https://huggingface.co/datasets/KamiBench/experiment-001-budget-boxed.tabularn<1K0 likes57 downloads2mo agoHugging Face23muahmed7338 /rl-experiment-rescue-lfstextn<1K0 likes52 downloads17d agoHugging Face24Debbevi /medical_dialogue_swe_experiment1Synthetic dataset of medical encounters in Swedish generated with Llama 3 70B. The conversations do not represent ideal communication but strive to capture conversations as they may happen in medical settings. Patients have a randomly sampled personality with assumed decorrelated normal distribution based on the psychometric data from Twenty-item, Table 3 in Ringwald, W. R. et al (2022). Psychometric evaluation of a Big Five personality state scale for intensive longitudinal studies.… See the full description on the dataset page: https://huggingface.co/datasets/Debbevi/medical_dialogue_swe_experiment1.text1K<n<10K0 likes48 downloads2y agoHugging Face25Praful932 /abmelt-experiments-exp_20260219_182250textn<1K0 likes42 downloads7mo agoHugging Face26Experimental-Orange /persona-belief-probestext10K<n<100K0 likes37 downloads4mo agoHugging Face27zifeng-ai /experimental-optimizationtextn<1K0 likes35 downloads4d agoHugging Face28Praful932 /abmelt-experiments-exp_20260216_181056textn<1K0 likes33 downloads7mo agoHugging Face29nyu-dice-lab /lm-eval-results-ChaoticNeutrals-Prima-LelantaclesV7-experimental-7b-private Dataset Card for Evaluation run of ChaoticNeutrals/Prima-LelantaclesV7-experimental-7b Dataset automatically created during the evaluation run of model ChaoticNeutrals/Prima-LelantaclesV7-experimental-7b The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-ChaoticNeutrals-Prima-LelantaclesV7-experimental-7b-private.tabular100K<n<1M0 likes29 downloads2y agoHugging Face30Praful932 /abmelt-experiments-exp_20260219_182408textn<1K0 likes27 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.