CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01dartbrains /localizer Dartbrains Localizer Dataset A subset of the Brainomics/Localizer functional MRI dataset, prepared for the Dartbrains neuroimaging course at Dartmouth College. Quick Start Load beta maps (recommended for most exercises) from datasets import load_dataset ds = load_dataset("dartbrains/localizer", "betas") img = ds[0]["nifti"] # nibabel.Nifti1Image subject = ds[0]["subject"] # "S01" condition = ds[0]["condition"] # "audio_computation"… See the full description on the dataset page: https://huggingface.co/datasets/dartbrains/localizer.imageimage-classificationn<1K1 likes756 downloads3mo agoHugging Face02OpenVoiceOS /ovos-localize-intents OpenVoiceOS Localize — Intent Classification Dataset Multilingual intent classification corpus exported from OpenVoiceOS/ovos-localize. Each row is a single expanded utterance labelled with the OVOS skill and intent file that produced it. Templates are fully expanded (bracket alternation resolved); {slot_name} placeholders from .intent files are kept verbatim so models can learn the slot-carrying pattern. Schema Column Description lang BCP-47 locale… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/ovos-localize-intents.text-classification100K<n<1M0 likes385 downloads2h agoHugging Face03mateoguaman /localized_narratives_trajectory_formatimage100K<n<1M0 likes302 downloads1y agoHugging Face04porhan /fmri-visual-localizerimagen<1K0 likes195 downloads18d agoHugging Face05localized-ft /selective-learning-benchmark-ip Selective Learning Benchmark Data: Inoculation Prompting This repository is an inoculation-prompting variant of localized-ft/selective-learning-benchmark. It bundles selective-learning task data in task_data_model_v1 JSONL format and prepends a subset-specific inoculation prompt as the system turn of every sft and validation example. The eval and control examples intentionally omit the prompt so evaluation measures learned behavior rather than direct prompt steering. Each task… See the full description on the dataset page: https://huggingface.co/datasets/localized-ft/selective-learning-benchmark-ip.text-generation0 likes95 downloads2mo agoHugging Face06sammlapp /Alberta_SBT_2016_REVI_Localized Red-eyed Vireo localized songs Creators: Sam Lapp (sam.lapp@pitt.edu) [1], Scott J. Wilson [2], Erin Bayne [3], and Justin Kitzes [1] Affiliations: [1] University of Pittsburgh, [2] Government of Alberta, [3] University of Alberta Version 1.1 Date Updated: 2026-09-22 DOI: not yet assigned General characteristics audio format: 10 second .FLAC clips starting 4 seconds before localized events dimensions localized: 2number of localization arrays: 13array geometry:… See the full description on the dataset page: https://huggingface.co/datasets/sammlapp/Alberta_SBT_2016_REVI_Localized.audion<1K0 likes95 downloads3d agoHugging Face07synthetic-code-training /func_localize_claude45_1457itext1K<n<10K0 likes86 downloads7d agoHugging Face08mateoguaman /cocoqa_localized_narratives cocoqa_localized_narratives Description Concatenated dataset of all of cocoqa and localized_narratives. Processing Parameters {} Dataset Configuration Train dataset: mixer: mateoguaman/cocoqa_trajectory_format: 1.0 mateoguaman/localized_narratives_trajectory_format: 1.0 split: train Validation dataset: mixer: mateoguaman/cocoqa_trajectory_format: 1.0 mateoguaman/localized_narratives_trajectory_format: 1.0 split: train… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/cocoqa_localized_narratives.image100K<n<1M0 likes74 downloads1y agoHugging Face09localized-ft /selective-learning-benchmark Selective Learning Benchmark Data This repository bundles selective-learning task data from Sunday, Srija, and Sultan in task_data_model_v1 JSONL format. Each task directory contains a manifest.json with contributor/source attribution, a capability description, an unintended-generalization description, split files, and row counts. Each Hugging Face config/subset is one dataset named as [type]-[name], with sft, validation, eval, and control splits where available. The type values… See the full description on the dataset page: https://huggingface.co/datasets/localized-ft/selective-learning-benchmark.text-generation0 likes70 downloads2mo agoHugging Face10synthetic-code-training /func_localize_claude47_1467itext1K<n<10K0 likes64 downloads7d agoHugging Face11elliot-mllm /localize-indoorgated Elliot Localize Indoor / 3D and depth Upstream training splits; known explicitly identified test/eval rows excluded. Cross-dataset benchmark overlap is not guaranteed. Published as a raw, manually gated release; source annotation caveats remain. Task views reuse original image archives. 2D coordinates are normalized 0–1000. Native 3D sidecars preserve original camera-space XYZ and camera calibration; they are not normalized to 0–1000. Within each query targets are sorted… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/localize-indoor.imageobject-detectionn<1K0 likes63 downloads12d agoHugging Face12synthetic-code-training /swe_bench_localize_sim_prompt_515itextn<1K0 likes61 downloads10mo agoHugging Face13AKCIT /localized_narrativesimage100K<n<1M1 likes56 downloads6mo agoHugging Face14nimapourjafar /mm_localized_narrativesimage100K<n<1M0 likes53 downloads2y agoHugging Face15elliot-mllm /localize-locany-voted-annotationsgated LocAny annotations Original annotations with saved Rex-Omni, Qwen, YOLO-E and SAM3 predictions where available. YOLO-E and SAM3 masks are stored as COCO RLE alongside their boxes. Images use the original media references. 49/49 original views uploaded (11,770,115 image records). Files: source/datasets/<dataset>/views/<view>/records.jsonl. Original fields are unchanged; _annotations holds model outputs. _cleaned_strict siblings contain the existing filtered annotations. These do… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/localize-locany-voted-annotations.0 likes51 downloads13d agoHugging Face16elliot-mllm /localize-guigated Elliot Localize GUI Upstream training splits; known explicitly identified test/eval rows excluded. Cross-dataset benchmark overlap is not guaranteed. Published as a raw, manually gated release; source annotation caveats remain. Task views reuse original image archives. Coordinates are normalized 0–1000. Within each query targets are sorted left-to-right then top-to-bottom. HF preview configs contain 10 examples per view, not the complete training split. Full training uses… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/localize-gui.imageobject-detectionn<1K0 likes49 downloads12d agoHugging Face17HuggingFaceM4 /LocalizedNarrativesLocalized Narratives, a new form of multimodal image annotations connecting vision and language. We ask annotators to describe an image with their voice while simultaneously hovering their mouse over the region they are describing. Since the voice and the mouse pointer are synchronized, we can localize every single word in the description. This dense visual grounding takes the form of a mouse trace segment per word and is unique to our data. We annotated 849k images with Localized Narratives: the whole COCO, Flickr30k, and ADE20K datasets, and 671k images of Open Images, all of which we make publicly available.image7 likes40 downloads4y agoHugging Face18synthetic-code-training /func_localize_claude47_min_file_explore_1467itext1K<n<10K0 likes35 downloads18d agoHugging Face19r2e-edits /localize-sympynew-gpt4o-v1text1K<n<10K0 likes34 downloads2y agoHugging Face20synthetic-code-training /ds3_swe_bench_localize_513itextn<1K0 likes32 downloads1y agoHugging Face21synthetic-code-training /func_localize_claude45_1457i_text300 func_localize_claude45_1457i_text300 Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 300 tokens (accepted band 225-375 tokens of the Qwen3 tokenizer, up to 3 rounds; 36 of 26194 turns missed the band and keep their original prose). Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are removed… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text300.text1K<n<10K0 likes30 downloads4d agoHugging Face22synthetic-code-training /func_localize_claude45_1457i_text2x func_localize_claude45_1457i_text2x Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 2 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15684 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 130 missed the band; both keep their original prose. Construction (shared by the _text0x/0.5x/2x/4x/8x… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text2x.text1K<n<10K0 likes30 downloads2d agoHugging Face23synthetic-code-training /func_localize_claude45_1457i_text0 func_localize_claude45_1457i_text0 Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: every assistant turn is its tool call only: the prose before the call is removed. Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are removed (trajectories contain only real tool calls); the system prompt, task, tool calls and tool results are byte-identical to the base. The rephraser saw only the current turn… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text0.text1K<n<10K0 likes29 downloads4d agoHugging Face24synthetic-code-training /func_localize_claude45_1457i_text4x func_localize_claude45_1457i_text4x Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 4 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15698 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 297 missed the band; both keep their original prose. Construction (shared by the _text0x/0.5x/2x/4x/8x… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text4x.text1K<n<10K0 likes29 downloads2d agoHugging Face25synthetic-code-training /func_localize_claude45_1457i_text100 func_localize_claude45_1457i_text100 Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 100 tokens (accepted band 75-125 tokens of the Qwen3 tokenizer, up to 3 rounds; 0 of 26194 turns missed the band and keep their original prose). Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are removed… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text100.text1K<n<10K0 likes28 downloads4d agoHugging Face26synthetic-code-training /func_localize_claude45_1457i_text0.5x func_localize_claude45_1457i_text0.5x Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 0.5 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15689 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 79 missed the band; both keep their original prose. Construction (shared by the… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text0.5x.text1K<n<10K0 likes28 downloads2d agoHugging Face27synthetic-code-training /func_localize_claude45_1457i_text8x func_localize_claude45_1457i_text8x Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to 8 times its own length in Qwen3 tokens (accepted band ±25 %, up to 3 rounds). Of 26194 turns, 15754 were not rephrased (empty prose, or a target outside 4-2000 tokens) and 360 missed the band; both keep their original prose. Construction (shared by the _text0x/0.5x/2x/4x/8x… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text8x.text1K<n<10K0 likes28 downloads2d agoHugging Face28synthetic-code-training /func_localize_claude45_1457i_text20 func_localize_claude45_1457i_text20 Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 20 tokens (accepted band 15-25 tokens of the Qwen3 tokenizer, up to 3 rounds; 178 of 26194 turns missed the band and keep their original prose). Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are removed (trajectories… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text20.text1K<n<10K0 likes27 downloads4d agoHugging Face29synthetic-code-training /func_localize_claude45_1457i_text50 func_localize_claude45_1457i_text50 Verbosity-ablation variant of synthetic-code-training/func_localize_claude45_1457i: the prose before every tool call is rewritten by nvidia/deepseek-ai/deepseek-v4-flash to about 50 tokens (accepted band 38-62 tokens of the Qwen3 tokenizer, up to 3 rounds; 8 of 26194 turns missed the band and keep their original prose). Construction (shared by all _text* siblings): think and task_tracker turns and their result turns are removed (trajectories… See the full description on the dataset page: https://huggingface.co/datasets/synthetic-code-training/func_localize_claude45_1457i_text50.0 likes26 downloads4d agoHugging Face30salma-remyx /vqasynth_cauldron_localized_narratives_100_fullimagen<1K0 likes25 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.