datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
localizer
Dartbrains Localizer Dataset
A subset of the Brainomics/Localizer functional MRI dataset, prepared for the Dartbrains neuroimaging course at Dartmouth College.
Quick Start
Load beta maps (recommended for most exercises)
from datasets import load_dataset
ds = load_dataset("dartbrains/localizer", "betas")
img = ds[0]["nifti"] # nibabel.Nifti1Image
subject = ds[0]["subject"] # "S01"
condition = ds[0]["condition"] # "audio_computation"… See the full description on the dataset page: https://huggingface.co/datasets/dartbrains/localizer.ovos-localize-intents
OpenVoiceOS Localize — Intent Classification Dataset
Multilingual intent classification corpus exported from
OpenVoiceOS/ovos-localize.
Each row is a single expanded utterance labelled with the OVOS skill and
intent file that produced it. Templates are fully expanded (bracket
alternation resolved); {slot_name} placeholders from .intent files are
kept verbatim so models can learn the slot-carrying pattern.
Schema
Column
Description
lang
BCP-47 locale… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/ovos-localize-intents.localized_narratives_trajectory_formatfmri-visual-localizerselective-learning-benchmark-ip
Selective Learning Benchmark Data: Inoculation Prompting
This repository is an inoculation-prompting variant of localized-ft/selective-learning-benchmark. It bundles selective-learning task data in task_data_model_v1 JSONL format and prepends a subset-specific inoculation prompt as the system turn of every sft and validation example. The eval and control examples intentionally omit the prompt so evaluation measures learned behavior rather than direct prompt steering.
Each task… See the full description on the dataset page: https://huggingface.co/datasets/localized-ft/selective-learning-benchmark-ip.Alberta_SBT_2016_REVI_Localized
Red-eyed Vireo localized songs
Creators: Sam Lapp (sam.lapp@pitt.edu) [1], Scott J. Wilson [2], Erin Bayne [3], and Justin Kitzes [1]
Affiliations: [1] University of Pittsburgh, [2] Government of Alberta, [3] University of Alberta
Version 1.1
Date Updated: 2026-09-22
DOI: not yet assigned
General characteristics
audio format: 10 second .FLAC clips starting 4 seconds before localized events
dimensions localized: 2number of localization arrays: 13array geometry:… See the full description on the dataset page: https://huggingface.co/datasets/sammlapp/Alberta_SBT_2016_REVI_Localized.func_localize_claude45_1457icocoqa_localized_narratives
cocoqa_localized_narratives
Description
Concatenated dataset of all of cocoqa and localized_narratives.
Processing Parameters
{}
Dataset Configuration
Train dataset:
mixer: mateoguaman/cocoqa_trajectory_format: 1.0
mateoguaman/localized_narratives_trajectory_format: 1.0
split: train
Validation dataset:
mixer: mateoguaman/cocoqa_trajectory_format: 1.0
mateoguaman/localized_narratives_trajectory_format: 1.0
split: train… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/cocoqa_localized_narratives.selective-learning-benchmark
Selective Learning Benchmark Data
This repository bundles selective-learning task data from Sunday, Srija, and Sultan in task_data_model_v1 JSONL format.
Each task directory contains a manifest.json with contributor/source attribution, a capability description, an unintended-generalization description, split files, and row counts.
Each Hugging Face config/subset is one dataset named as [type]-[name], with sft, validation, eval, and control splits where available. The type values… See the full description on the dataset page: https://huggingface.co/datasets/localized-ft/selective-learning-benchmark.func_localize_claude47_1467ilocalize-indoor
Elliot Localize Indoor / 3D and depth
Upstream training splits; known explicitly identified test/eval rows excluded. Cross-dataset benchmark overlap is not guaranteed. Published as a raw, manually gated release; source annotation caveats remain.
Task views reuse original image archives. 2D coordinates are normalized 0–1000. Native 3D sidecars preserve original camera-space XYZ and camera calibration; they are not normalized to 0–1000. Within each query targets are sorted… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/localize-indoor.swe_bench_localize_sim_prompt_515ilocalized_narrativesmm_localized_narrativeslocalize-locany-voted-annotations
LocAny annotations
Original annotations with saved Rex-Omni, Qwen, YOLO-E and SAM3 predictions where available. YOLO-E and SAM3 masks are stored as COCO RLE alongside their boxes. Images use the original media references.
49/49 original views uploaded (11,770,115 image records).
Files: source/datasets/<dataset>/views/<view>/records.jsonl. Original fields are unchanged; _annotations holds model outputs. _cleaned_strict siblings contain the existing filtered annotations. These do… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/localize-locany-voted-annotations.localize-gui
Elliot Localize GUI
Upstream training splits; known explicitly identified test/eval rows excluded. Cross-dataset benchmark overlap is not guaranteed. Published as a raw, manually gated release; source annotation caveats remain.
Task views reuse original image archives. Coordinates are normalized 0–1000. Within each query targets are sorted left-to-right then top-to-bottom.
HF preview configs contain 10 examples per view, not the complete training split. Full training uses… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/localize-gui.LocalizedNarrativesLocalized Narratives, a new form of multimodal image annotations connecting vision and language.
We ask annotators to describe an image with their voice while simultaneously hovering their mouse over the region they are describing.
Since the voice and the mouse pointer are synchronized, we can localize every single word in the description.
This dense visual grounding takes the form of a mouse trace segment per word and is unique to our data.
We annotated 849k images with Localized Narratives: the whole COCO, Flickr30k, and ADE20K datasets, and 671k images of Open Images, all of which we make publicly available.func_localize_claude47_min_file_explore_1467ilocalize-sympynew-gpt4o-v1ds3_swe_bench_localize_513ifunc_localize_claude45_1457i_text300
func_localize_claude45_1457i_text300
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 300 tokens (accepted band 225-375 tokens of the Qwen3 tokenizer, up to 3 rounds; 36 of 26194 turns missed the band and keep their original prose).
Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are
removed… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text300.func_localize_claude45_1457i_text2x
func_localize_claude45_1457i_text2x
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 2 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15684 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 130 missed the band; both keep their original prose.
Construction (shared by the _text0x/0.5x/2x/4x/8x… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text2x.func_localize_claude45_1457i_text0
func_localize_claude45_1457i_text0
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: every assistant turn is its tool call only: the prose before the call is removed.
Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are
removed (trajectories contain only real tool calls); the system prompt, task, tool calls and tool results are
byte-identical to the base. The rephraser saw only the current turn… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text0.func_localize_claude45_1457i_text4x
func_localize_claude45_1457i_text4x
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 4 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15698 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 297 missed the band; both keep their original prose.
Construction (shared by the _text0x/0.5x/2x/4x/8x… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text4x.func_localize_claude45_1457i_text100
func_localize_claude45_1457i_text100
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 100 tokens (accepted band 75-125 tokens of the Qwen3 tokenizer, up to 3 rounds; 0 of 26194 turns missed the band and keep their original prose).
Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are
removed… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text100.func_localize_claude45_1457i_text0.5x
func_localize_claude45_1457i_text0.5x
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 0.5 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15689 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 79 missed the band; both keep their original prose.
Construction (shared by the… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text0.5x.func_localize_claude45_1457i_text8x
func_localize_claude45_1457i_text8x
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 8 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15754 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 360 missed the band; both keep their original prose.
Construction (shared by the _text0x/0.5x/2x/4x/8x… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text8x.func_localize_claude45_1457i_text20
func_localize_claude45_1457i_text20
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 20 tokens (accepted band 15-25 tokens of the Qwen3 tokenizer, up to 3 rounds; 178 of 26194 turns missed the band and keep their original prose).
Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are
removed (trajectories… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text20.func_localize_claude45_1457i_text50
func_localize_claude45_1457i_text50
Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 50 tokens (accepted band 38-62 tokens of the Qwen3 tokenizer, up to 3 rounds; 8 of 26194 turns missed the band and keep their original prose).
Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are
removed (trajectories… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text50.vqasynth_cauldron_localized_narratives_100_full
