datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
BlendedMVS_processedBlenderLore
The target scope is 22,360 video-associated Blender project instances, not 22,360 distinct tutorial videos. Uploads are in progress, so the currently published files may be a subset of this target. The 44 biomedical project instances and one software-bundled Dome template are excluded.
Data Structure
The dataset is organized as a collection of sample-level directories under assets/. Each directory corresponds to one Blender creation task and follows the structure below:… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/BlenderLore.bleach
Bangumi Image Base of Bleach
This is the image base of bangumi Bleach, we detected 181 characters, 30903 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is the characters' preview:… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/bleach.Hoyoverse_Character_ModelsUltraEdit_500k
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/BleachNick/UltraEdit_500k.cua-blenderBlenderBench
BlenderBench Dataset
Dataset Description
BlenderBench is a comprehensive benchmark dataset for evaluating models on 3D scene editing tasks in Blender. The dataset challenges agents to understand visual differences between initial and target scenes, then generate appropriate Blender Python code to transform the initial scene to match the target.
Key Features
27 instances across 3 difficulty levels
Multi-modal: Combines visual… See the full description on the dataset page: https://huggingface.co/datasets/DietCoke4671/BlenderBench.NOAA-PIFSC-ESD-CORAL-Bleaching-Dataset
Dataset Card for NOAA-ESD-CORAL-Bleaching Classification Dataset v1
Overview
For the development of machine learning models to classify coral health, specifically identifying healthy hard coral (CORAL) and bleached hard coral (CORAL_BL).This dataset contains underwater imagery collected by NOAA's Ecosystem Sciences Division (ESD) and other benthic surveys.
Labels
Label
Name
Functional Group
CORAL
Healthy Hard Coral
Hard Coral
CORAL_BL… See the full description on the dataset page: https://huggingface.co/datasets/De129472/NOAA-PIFSC-ESD-CORAL-Bleaching-Dataset.UltraEdit_Region_Based_100k
Bibtex citation
@misc{zhao2024ultraeditinstructionbasedfinegrainedimage,
title={UltraEdit: Instruction-based Fine-Grained Image Editing at Scale},
author={Haozhe Zhao and Xiaojian Ma and Liang Chen and Shuzheng Si and Rujie Wu and Kaikai An and Peiyu Yu and Minjia Zhang and Qing Li and Baobao Chang},
year={2024},
eprint={2407.05282},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2407.05282},
}
BlenderRAG
BlenderRAG Dataset
A dataset for 3D scene and object generation research. Each sample pairs a Blender Python script that procedurally generates a 3D object with a rendered preview image and a natural-language description.
Dataset Summary
The dataset is organized into two top-level scenes — indoor and outdoor — each containing a collection of objects. Every object is represented by three aligned modalities:
File
Modality
Purpose
code_n.py
Python (Blender API)… See the full description on the dataset page: https://huggingface.co/datasets/MaxRondelli/BlenderRAG.blenderbench-direct-results
BlenderBench Direct reproduction artifacts
This dataset preserves artifacts and provenance for an independent, community-run reproduction of the public BlenderBench task set. It is not an official Blender Foundation product, official BlenderBench submission, or leaderboard result.
Experiment
Dataset: DietCoke4671/BlenderBench revision 203e4d325e9438ca55b29bdfc4f6a90842d74e68
Attribution: DietCoke4671 and contributors, CC BY 4.0
Generation model: gpt-6-astra
Codex… See the full description on the dataset page: https://huggingface.co/datasets/michaelgold/blenderbench-direct-results.Hoyoverse_MapsNItems_BlenderBLEnD-Vis
BLEnD-Vis
BLEnD-Vis is a benchmark for evaluating vision-language models (VLMs) on culturally grounded multiple-choice questions, including a text-only setting and a visual setting with generated images.
Paper: https://arxiv.org/abs/2510.11178
Dataset repo: https://huggingface.co/datasets/Incomple/BLEnD-Vis
Code: https://github.com/Social-AI-Studio/BLEnD-Vis
Source
BLEnD-Vis is derived from the BLEnD dataset on Hugging Face (nayeon212/BLEnD).
What is in… See the full description on the dataset page: https://huggingface.co/datasets/Incomple/BLEnD-Vis.NOAA-PIFSC-ESD-CORAL-Bleaching-Dataset
Dataset Card for NOAA-ESD-CORAL-Bleaching Classification Dataset v1
Overview
For the development of machine learning models to classify coral health, specifically identifying healthy hard coral (CORAL) and bleached hard coral (CORAL_BL).This dataset contains underwater imagery collected by NOAA's Ecosystem Sciences Division (ESD) and other benthic surveys.
Labels
Label
Name
Functional Group
CORAL
Healthy Hard Coral
Hard Coral
CORAL_BL
Bleached… See the full description on the dataset page: https://huggingface.co/datasets/NMFS-OSI/NOAA-PIFSC-ESD-CORAL-Bleaching-Dataset.textures-for-blenderBlenderkit_retail_bench_mini
Blenderkit Retail Bench Mini
Mini subset of the Blenderkit retail 3D-text benchmark (131 assets, ~4.8 GB).
Layout
Path
Description
Source GLB assets
Conditioning renders
Mesh dumps
PBR dumps
Dual-grid views (1024)
PBR voxels
Shape latents
PBR / texture latents
Per-asset metadata
Aggregate stats
Quick start
Or with mirror:
NOAA-PIFSC-ESD-CORAL-Bleaching-Dataset
Dataset Card for NOAA-ESD-CORAL-Bleaching Classification Dataset v1
Overview
For the development of machine learning models to classify coral health, specifically identifying healthy hard coral (CORAL) and bleached hard coral (CORAL_BL).This dataset contains underwater imagery collected by NOAA's Ecosystem Sciences Division (ESD) and other benthic surveys.
Labels
Label
Name
Functional Group
CORAL
Healthy Hard Coral
Hard Coral
CORAL_BL… See the full description on the dataset page: https://huggingface.co/datasets/Kshoarya-8/NOAA-PIFSC-ESD-CORAL-Bleaching-Dataset.blender-open-movies-mediatransportgs-blender-scenes
TransportGS Blender Scenes
One continuous RGB input stream at 2× speed, with no cut (sequence_000).
These synthetic RGB-D sequences study streaming Gaussian reconstruction when an object moves outside the camera's view. The camera first observes the object, turns away during the move, then returns for a partial revisit. A method must use the observation history to update the object at its new location and clear its old location. no_change sequences test whether it avoids… See the full description on the dataset page: https://huggingface.co/datasets/pengyue-polaron/transportgs-blender-scenes.BlenderGaze
BlenderGaze
The BlenderGaze dataset is specifically created to further investigate Visual Perspective Taking (VPT) in Vision Language Models (VLMs). This dataset extends the previously introduced Isle dataset, primarily expanding in terms of size rather than diversity.
BlenderGaze consists of two complementary subsets:
BlenderGaze-Isolated: Contains isolated scenes featuring a humanoid figure and a red box. These scenes are explicitly designed to isolate and test the basic VPT… See the full description on the dataset page: https://huggingface.co/datasets/Gracjan/BlenderGaze.blends
Bangumi Image Base of Blend S
This is the image base of bangumi Blend S, we detected 16 characters, 1863 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is the characters' preview:… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/blends.blendedmvsBlender⚠️ Note concernant les droits d'auteur : J'ai fait de mon mieux pour m'assurer que
tout ce contenu est original ou libre de droits. Cependant, vu le volume de fichiers,
il est possible qu'une texture, une image ou un objet appartenant à un autre créateur
s'y soit glissé. Si vous reconnaissez l'une de vos œuvres et que vous souhaitez
être crédité ou demander le retrait du fichier concerné, n'hésitez pas à me contacter
en message privé. Je ferai le nécessaire immédiatement !
WuWa_MapsSunBank-blender-outputsBlenderCAD2bleedingheart-pretrain-10MBleedingheart Pretrain Dataset
A collaboration between Kaleido and Newstar
We collected all the datasets we could find that are in Tagalog or any other Philippine dialect and put them in this repository.
This data will be used to train the Bleedingheart model.
Bleeding Heart is a stunning bird native to the island of Luzon in the Philippines. It is a medium-sized ground dove with a distinctive red patch of feathers on its chest, which gives it its name. The male's red patch is… See the full description on the dataset page: https://huggingface.co/datasets/NewstaR/bleedingheart-pretrain-10M.CIFAR100-Blended-20pct-Backdoor-ExclNaturalCraftBenchgemma-ft-dataset
Dataset Card for "gemma-ft-dataset"
More Information needed
