datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mid-space
MID-Space: Aligning Diverse Communities’ Needs to Inclusive Public Spaces
A new version of the dataset will be released soon, incorporating user identity markers and expanded annotations.
LIVS PAPER
Click below to see more:
Overview
The MID-Space dataset is designed to align AI-generated visualizations of urban public spaces with the preferences of diverse and marginalized communities in Montreal. It includes textual prompts, Stable Diffusion… See the full description on the dataset page: https://huggingface.co/datasets/mila-ai4h/mid-space.SpatialEval
🤔 About SpatialEval
SpatialEval is a comprehensive benchmark for evaluating spatial intelligence in LLMs and VLMs across four key dimensions:
Spatial relationships
Positional understanding
Object counting
Navigation
Benchmark Tasks
Spatial-Map: Understanding spatial relationships between objects in map-based scenarios
Maze-Nav: Testing navigation through complex environments
Spatial-Grid: Evaluating spatial reasoning within structured environments
Spatial-Real:… See the full description on the dataset page: https://huggingface.co/datasets/MilaWang/SpatialEval.milady
Dataset Card for "milady"
More Information needed
Milady-Avatar-Dataset
Milady Avatar Dataset
Training dataset for the Milady Avatar Adapter, mapping LLM emotional activations
to Milady NFT-style visual descriptions.
Contents
reference_images/: 200 Milady NFT reference images (1000×1250 PNG)
metadata.json: Emotion assignments and descriptions for each image
training_data.pt: Pre-computed activations and target embeddings
Structure
200 images assigned to 20 emotion categories (10 each):
happy, sad, angry, surprised, scared… See the full description on the dataset page: https://huggingface.co/datasets/Alogotron/Milady-Avatar-Dataset.MuSP-Bench
MuSP-Bench
MuSP-Bench is a 490-question benchmark for musical score understanding,
performance listening, and combined score-performance reasoning.
Contents
data/questions.csv: all 490 questions, accepted answers, and the
response contract for each.
inputs/pdf/without_context/: one context-removed PDF per piece.
inputs/images/: rendered score-page images for every piece.
inputs/abc/: one ABC score per piece.
inputs/abc_plus_midi/: one aligned ABC+MIDI… See the full description on the dataset page: https://huggingface.co/datasets/milan477/MuSP-Bench.Brain-MRI-Images-for-Brain-Tumor-Detection
Brain Tumor Detection | Vision Transformer 99%
Click -> Kaggle
task_categories:
- image-classification
- image-segmentation
tags:
- 'brain '
- MRI
- brain-MRI-images
- Tumor
Intel-Image-Classificationmilady
Milady
Milady Maker is a collection of 10,000 generative pfpNFT's in a neochibi aesthetic inspired by street style tribes.
5-Flower-Types-Classification-Datasetnaruto-blip-captions
Dataset Card for Naruto BLIP captions
Dataset used to train TBD.
The original images were obtained from narutopedia.com and captioned with the pre-trained BLIP model.
For each row the dataset contains image and text keys. image is a varying size PIL jpeg, and text is the accompanying text caption. Only a train split is provided.
Example stable diffusion outputs
"Bill Gates with a hoodie", "John Oliver with Naruto style", "Hello Kitty with Naruto style", "Lebron… See the full description on the dataset page: https://huggingface.co/datasets/Milabench/naruto-blip-captions.DeepFa1Rice-Image-Datasetspatial_traveres_results_docent-val-Molmo2-8B-spatial_no_refdocent-val-Qwen3-VL-8B-Instruct-boundary_no_refspatial_traveres_results_docent-val-Qwen3-VL-8B-Instruct-spatial-w-semantic_no_refAstroStarDetectSTL-Fashion-Product-Testspatial_traveres_results_docent-val-Molmo2-8B-semantic_no_refdocent-val-Molmo2-8B-boundary-w-text_no_refspatial_traveres_results_docent-val-Qwen3-VL-8B-Instruct-semantic_no_refspatial_traveres_results_docent-val-Molmo2-8B-spatial-w-semantic_no_refdocent-val-InternVL2_5-8B-bbox-w-text_no_refdocent-val-InternVL2_5-8B-boundary_no_refdocent-val-Molmo2-8B-boundary_no_refspatial_traveres_results_docent-val-Qwen3-VL-8B-Instruct-for-to-background_no_refSTL-Fashion-ProductDeepfa1_trainquickdraw-coarsedocent-val-Molmo2-8B-bbox_no_refdocent-val-Molmo2-8B-bbox-w-text_no_ref
