datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Egocentric-100K
Egocentric-100K is the largest dataset of manual labor. You can visualize the dataset here.
Egocentric-100K is state-of-the-art in hand visibility and active manipulation density compared to previous in-the-wild egocentric datasets. The complete 30,000 frame evaluation set is available at Egocentric-100K-Evaluation.
Dataset Statistics
Attribute
Value
Total Hours
100,405
Total Frames
10.8 billion
Video Clips
2,010,759
Median Clip Length
180.0 seconds
Mean… See the full description on the dataset page: https://huggingface.co/datasets/builddotai/Egocentric-100K.TimeLens-100K
TimeLens-100K
📑 Paper | 💻 Code | 🏠 Project Page | 🤗 Model & Data
✨ Dataset Description
TimeLens-100K is a large-scale, diverse, and high-quality training dataset for video temporal grounding. It was proposed in our paper TimeLens: Rethinking Video Temporal Grounding with Multimodal LLMs and used for training TimeLens models. The annotation process was conducted using an automated pipeline powered by Gemini-2.5-Pro.
📊 Dataset Statistics
Total Videos:… See the full description on the dataset page: https://huggingface.co/datasets/TencentARC/TimeLens-100K.RevealLayer-100K
RevealLayer Open Dataset
RevealLayer Open is the open-source dataset accompanying RevealLayer: Disentangling Hidden and Visible Layers via Occlusion-Aware Image Decomposition.
Paper: https://arxiv.org/html/2605.11818v1 Accepted by ICML 2026
RevealLayer studies box-guided layered image decomposition for natural images. Given an RGB image and instance bounding boxes, the task is to decompose the scene into a clean background and object-level foreground layers, where each… See the full description on the dataset page: https://huggingface.co/datasets/qihoo360/RevealLayer-100K.cc3m-subset-100kMieDB-100k
MieDB-100k: A Comprehensive Dataset for Medical Image Editing
📄 Introduction
MieDB-100k is a large-scale, high-quality and diverse dataset for text-guided medical image editing,
which includes 104,267 editing data, covering 63 distinct editing targets and 10 diverse medical image modalities.
We categorize editing tasks into three types: Perception, Modification and Transformation, which consider both model's intrinsic understanding and generation abilities on medical… See the full description on the dataset page: https://huggingface.co/datasets/Laiyf/MieDB-100k.Nano3D-Edit-100k
Nano3D-Edit-100k
This dataset is the official data release for Nano3D, a training-free framework for precise and coherent 3D object editing without masks.
Paper: Nano3D: A Training-Free Approach for Efficient 3D Editing Without MasksProject Page: https://jamesyjl.github.io/Nano3D/
Nano3D integrates FlowEdit into TRELLIS to perform localized 3D edits guided by front-view renderings, and introduces Voxel/Slat-Merge strategies to preserve structural consistency between edited and… See the full description on the dataset page: https://huggingface.co/datasets/yejunliang23/Nano3D-Edit-100k.Movielens-100k-framesgradienttile-100kA large-scale dataset of 100K procedurally generated animated GIFs featuring mathematical patterns, fractals, and artistic visualizations.
aesthetic-100kThis repo host first 100k of anime aesthetic dataset, intended for post training T2I models for anime image generation. As aesthetic is not well defined I decided to take liberty with it and included images that are considered "hard" to text-to-image generating system.
If you see no dataset files then I haven't done with it yet
tic_tac_toe_100kwds-mops-kitchen-100kThis is the anonymous Dataset upload for Multi-Objective Photoreal Simulation (MOPS) Dataset [under review].
Official code available at: https://github.com/mopsneurips26-code/mops-submission
