CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01evalstate /transformers-merge-experimentstabularn<1K3 likes1.8k downloads5mo agoHugging Face02aimosprite /prompt-swap-mixed12-5xlr-e1-mxfp4-mergedtabularn<1K0 likes530 downloads6mo agoHugging Face03aimosprite /prompt-swap-mixed12-5xlr-e2-mxfp4-mergedtabularn<1K0 likes515 downloads6mo agoHugging Face04NuTonic /sat-vl-sft-postprocessed-merged-v1 Dataset Summary NuTonic/sat-bbox-metadata-sft-v1 is a metadata-first, procedural VLM SFT dataset built from an existing “sat-bbox” style dataset tree (Sentinel‑2 chips + per-tile JSON metadata sidecars, optionally paired Mapbox stills). The goal is to create high-signal, production-shaped supervision for multimodal chat models: Captioning for satellite chips Grounding (bounding boxes in normalized coordinates) for land-cover regions Class-focused captions and absence checks for… See the full description on the dataset page: https://huggingface.co/datasets/NuTonic/sat-vl-sft-postprocessed-merged-v1.imagetext-generation100K<n<1M0 likes502 downloads5mo agoHugging Face05aimosprite /prompt-swap-medium12-e2-mxfp4-mergedtabularn<1K0 likes343 downloads6mo agoHugging Face06survivi /grad_clip0.28_mergedtext100K<n<1M0 likes342 downloads1y agoHugging Face07w4nn4b3M4ST3R /raw-mergedtabular100K<n<1M0 likes288 downloads2mo agoHugging Face08Chinese-Vicuna /guanaco_belle_merge_v1.0Thanks for Guanaco Dataset and Belle Dataset This dataset was created by merging the above two datasets in a certain format so that they can be used for training our code Chinese-Vicuna text100K<n<1M101 likes226 downloads3y agoHugging Face09nyu-dice-lab /lm-eval-results-alnrg2arg-blockchainlabs_7B_merged_test2_4-private Dataset Card for Evaluation run of alnrg2arg/blockchainlabs_7B_merged_test2_4 Dataset automatically created during the evaluation run of model alnrg2arg/blockchainlabs_7B_merged_test2_4 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-alnrg2arg-blockchainlabs_7B_merged_test2_4-private.tabular100K<n<1M0 likes183 downloads2y agoHugging Face10michaeldinzinger /merged-trecdltexttext-retrieval1M<n<10M0 likes153 downloads1y agoHugging Face11Cross-Mergeability /merge-accuracy merge-accuracy — does aligning a merge improve DOWNSTREAM ACCURACY? The mergeability line of work measures merge obstruction in nats/token. This dataset supplies the missing axis: task accuracy, on real released models, for the merge recipe practitioners actually run — the chat-vector recipe theta_new = theta_fork + lambda * ( theta_instruct - theta_base ) with meta-llama/Llama-3.1-8B, its official Instruct release, and three community continued-pretrained language forks… See the full description on the dataset page: https://huggingface.co/datasets/Cross-Mergeability/merge-accuracy.textn<1K0 likes151 downloads1mo agoHugging Face12Sugita-daichi /LoRA-Merge-Imagesimageimage-to-image100K<n<1M0 likes127 downloads5mo agoHugging Face13ambrosfitz /textbook-openstax-yawp-mergetext1K<n<10K0 likes117 downloads3y agoHugging Face14khtsly /cache_stage1_mergedtabularn<1K0 likes107 downloads16d agoHugging Face15nyu-dice-lab /lm-eval-results-chlee10-T3Q-Merge-Mistral7B-private Dataset Card for Evaluation run of chlee10/T3Q-Merge-Mistral7B Dataset automatically created during the evaluation run of model chlee10/T3Q-Merge-Mistral7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chlee10-T3Q-Merge-Mistral7B-private.tabular100K<n<1M0 likes99 downloads2y agoHugging Face16nyu-dice-lab /lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-private Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-private.tabular100K<n<1M0 likes97 downloads2y agoHugging Face17WithinUsAI /fable_5_distillation_merged_cleaned_25k Claude Fable 5 Distillation Dataset 25,719 high-quality distilled examples for training LLMs to mimic Claude Fable 5's reasoning style — featuring multi-step chain-of-thought with <think> tags across 23+ technical domains. This dataset captures the distinctive reasoning patterns of Claude Fable 5 (Anthropic's Mythos-class model released June 2026): systematic decomposition, first-principles analysis, self-verification, alternative consideration, and synthesis.… See the full description on the dataset page: https://huggingface.co/datasets/WithinUsAI/fable_5_distillation_merged_cleaned_25k.text10K<n<100K6 likes91 downloads3mo agoHugging Face18professorsynapse /claudesidian-behaviors-merged Claudesidian Merged Behavioral Dataset Dataset Description This dataset contains 1,852 synthetic training examples demonstrating 8 different behavioral patterns for training language models to use the Claudesidian-MCP toolset effectively with Obsidian vaults. The dataset is specifically formatted for KTO (Kahneman-Tversky Optimization) preference learning with properly interleaved positive and negative examples. Behavioral Categories This dataset includes… See the full description on the dataset page: https://huggingface.co/datasets/professorsynapse/claudesidian-behaviors-merged.texttext-generation1K<n<10K0 likes90 downloads10mo agoHugging Face19nyu-dice-lab /lm-eval-results-MiniMoog-Mergerix-7b-v0.5-private Dataset Card for Evaluation run of MiniMoog/Mergerix-7b-v0.5 Dataset automatically created during the evaluation run of model MiniMoog/Mergerix-7b-v0.5 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-MiniMoog-Mergerix-7b-v0.5-private.tabular100K<n<1M0 likes71 downloads2y agoHugging Face20nyu-dice-lab /lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test-private Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test-private.tabular100K<n<1M0 likes69 downloads2y agoHugging Face21nyu-dice-lab /lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-v2-private Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-v2 Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-v2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-v2-private.tabular100K<n<1M0 likes69 downloads2y agoHugging Face22nyu-dice-lab /lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-private Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-private.tabular100K<n<1M0 likes68 downloads2y agoHugging Face23nyu-dice-lab /lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2-private Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2 Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2-private.tabular100K<n<1M0 likes66 downloads2y agoHugging Face24prashantss1404 /Matplotlib_Seaborn_merged_prompt_completion_10ktext1K<n<10K0 likes66 downloads1y agoHugging Face25nyu-dice-lab /lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3-private Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3 Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3-private.tabular100K<n<1M0 likes65 downloads2y agoHugging Face26Lazycuber /Evol-instruct-mergejust merging a few evol instruct datasets the pros made text100K<n<1M1 likes63 downloads3y agoHugging Face27Lyric1010 /numina-0.02B-drop-merge-llama Dataset: numina-0.02B-drop-merge-llama This dataset was uploaded from /mnt/yulan_pretrain/mount/data_final_train_llama3/numina-0.02B-drop-merge/stage_1. textn<1K0 likes61 downloads10mo agoHugging Face28Lyric1010 /numina-0.2B-drop-entropy50-merge Dataset: numina-0.2B-drop-entropy50-merge This dataset was uploaded from /mnt/yulan_pretrain/mount/data_final_train_qwen3/numina-0.2B-drop-entropy50-merge/stage_1. textn<1K0 likes60 downloads10mo agoHugging Face29flatlander1024 /math_merged_cot_sol_pair_mixedPair Type Breakdown: Correct-Incorrect (C-I) Pairs: 5500 C-I with correct first ('[1]'): 2750 C-I with correct second ('[2]'): 2750 Correct-Correct (C-C) Pairs (Target: 2750, Max Diff: 150): 2750 C-C pairs from 'all_correct' problems: 906 Incorrect-Incorrect (I-I) Pairs (Target: 2750, Max Diff: 150): 2750 I-I pairs from 'all_incorrect' problems: 1156 text10K<n<100K0 likes59 downloads1y agoHugging Face30MERGE-Group /PATH-VQAtext10K<n<100K0 likes56 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.