datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
prompt-swap-mixed12-5xlr-e1-mxfp4-mergedprompt-swap-mixed12-5xlr-e2-mxfp4-mergedsat-vl-sft-postprocessed-merged-v1
Dataset Summary
NuTonic/sat-bbox-metadata-sft-v1 is a metadata-first, procedural VLM SFT dataset built from an existing “sat-bbox” style dataset tree (Sentinel‑2 chips + per-tile JSON metadata sidecars, optionally paired Mapbox stills).
The goal is to create high-signal, production-shaped supervision for multimodal chat models:
Captioning for satellite chips
Grounding (bounding boxes in normalized coordinates) for land-cover regions
Class-focused captions and absence checks for… See the full description on the dataset page: https://huggingface.co/datasets/NuTonic/sat-vl-sft-postprocessed-merged-v1.prompt-swap-medium12-e2-mxfp4-mergedgrad_clip0.28_mergedraw-mergedlm-eval-results-alnrg2arg-blockchainlabs_7B_merged_test2_4-private
Dataset Card for Evaluation run of alnrg2arg/blockchainlabs_7B_merged_test2_4
Dataset automatically created during the evaluation run of model alnrg2arg/blockchainlabs_7B_merged_test2_4
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-alnrg2arg-blockchainlabs_7B_merged_test2_4-private.merged-trecdlbanana-merged
Banana-Merged
A synthetic multi-page visual question answering dataset with hard negatives, designed for fine-tuning visual document retrievers like ColFlor and ColPali.
Dataset Summary
Banana-Merged contains 1,100 training samples and 10,054 images (positive pages + hard negative variants). Each sample pairs a multi-page analytical query with a set of document images that collectively contain the answer, plus one or more hard negative documents that look visually and… See the full description on the dataset page: https://huggingface.co/datasets/vkehfdl1/banana-merged.cache_stage1_mergedfable_5_distillation_merged_cleaned_25k
Claude Fable 5 Distillation Dataset
25,719 high-quality distilled examples for training LLMs to mimic Claude Fable 5's reasoning style — featuring multi-step chain-of-thought with <think> tags across 23+ technical domains.
This dataset captures the distinctive reasoning patterns of Claude Fable 5 (Anthropic's Mythos-class model released June 2026): systematic decomposition, first-principles analysis, self-verification, alternative consideration, and synthesis.… See the full description on the dataset page: https://huggingface.co/datasets/WithinUsAI/fable_5_distillation_merged_cleaned_25k.claudesidian-behaviors-merged
Claudesidian Merged Behavioral Dataset
Dataset Description
This dataset contains 1,852 synthetic training examples demonstrating 8 different behavioral patterns for training language models to use the Claudesidian-MCP toolset effectively with Obsidian vaults.
The dataset is specifically formatted for KTO (Kahneman-Tversky Optimization) preference learning with properly interleaved positive and negative examples.
Behavioral Categories
This dataset includes… See the full description on the dataset page: https://huggingface.co/datasets/professorsynapse/claudesidian-behaviors-merged.Matplotlib_Seaborn_merged_prompt_completion_10kmath_merged_cot_sol_pair_mixedPair Type Breakdown:
Correct-Incorrect (C-I) Pairs: 5500
C-I with correct first ('[1]'): 2750
C-I with correct second ('[2]'): 2750
Correct-Correct (C-C) Pairs (Target: 2750, Max Diff: 150): 2750
C-C pairs from 'all_correct' problems: 906
Incorrect-Incorrect (I-I) Pairs (Target: 2750, Max Diff: 150): 2750
I-I pairs from 'all_incorrect' problems: 1156
prompt-swap-medium12-e1-mxfp4-mergedmath_mergedTraining dataset contains aime (excluding 2024), math/train, math/test, openai_math_splits/train, and KbsdJames/Omni-MATH/test. Total of 17521 lines of unique problems.
Testing dataset contains aime_24 and math500 (i.e. openai_math_splits/test). Total of 530 lines of unique problems.
opus-agent-merged-v2Cleaned-sharegpt_Merged-Opus-33159-ShareGPTDans-DiscountModels__mistral-7b-test-merged-details
Dataset Card for Evaluation run of Dans-DiscountModels/mistral-7b-test-merged
Dataset automatically created during the evaluation run of model Dans-DiscountModels/mistral-7b-test-merged
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Dans-DiscountModels__mistral-7b-test-merged-details.openthoughts_merged_think_39k
openthoughts_merged_think_39k
Merged think-format SFT dataset (ShareGPT-style: system + conversations with
from/value), 39,874 examples, for OLMo SFT.
Composition (concatenation of two decontaminated think-format sources):
open-thoughts114k_math_20k_decontam_think — 20,000 examples sampled from
OpenThoughts-114k math, decontaminated against the OpenThoughts3 set below.
openthoughts3_math_decontam_resp_lt8192_think — 19,874 examples from… See the full description on the dataset page: https://huggingface.co/datasets/pre-to-post-olmo/openthoughts_merged_think_39k.FlofloB__40k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit-details
Dataset Card for Evaluation run of FlofloB/40k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit
Dataset automatically created during the evaluation run of model FlofloB/40k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FlofloB__40k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit-details.Merged_Nepali_Health_FAQ_Dataset
Merged Nepali Health FAQ Dataset (Multi-Source)
Overview
This dataset (merged_nepali_sharegpt_after_removing_11_groups_keep_ids.jsonl) is a merged collection of 65 instruction-following conversation pairs in Nepali, combining Q&A content from 8 distinct real-world Nepali health institutions and organizations into a single ShareGPT-style file. Each record is a single-turn human↔gpt exchange: a Nepali-language question followed by a factual Nepali-language answer.… See the full description on the dataset page: https://huggingface.co/datasets/sabin1234/Merged_Nepali_Health_FAQ_Dataset.FlofloB__100k_fineweb_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit-details
Dataset Card for Evaluation run of FlofloB/100k_fineweb_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit
Dataset automatically created during the evaluation run of model FlofloB/100k_fineweb_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FlofloB__100k_fineweb_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bit-details.merged-admin-def-dataset-16kopus-claude-mergedgemini-3.1-opus-4.6-reasoning-merged Merged from
https://huggingface.co/datasets/Roman1111111/gemini-3.1-pro-hard-high-reasoning
https://huggingface.co/datasets/reedmayhew/gemini-3.1-pro-2048-reasoning-1100x
https://huggingface.co/datasets/crownelius/Opus-4.6-Reasoning-3300x
WizardLM_evol_instruct_V2_196k_unfiltered_merged_splittranscripts-mergedPyThagoreans-Merged
PyThagoreans Dataset
Overview
The PyThagoreans dataset is a comprehensive collection of math problems and their solutions, designed to assist in learning and practicing mathematical problem-solving. This dataset includes a variety of problems, expected answers, and predicted answers, making it a valuable resource for students, educators, and researchers.
Dataset Details
Modalities
Text: The dataset primarily contains text data, including math… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/PyThagoreans-Merged.math_merged_cot_solA dataset consists problems from flatlander1024/math_merged and cot solutions generated by Llama-3.1-8b-Instruct. The is_correct label indicates whether the solution is correct or not.
Number of lines: 13864, Overall correct rate: 57.3%
