datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OneThinker-train-data
OneThinker-600k Training Data
This repository contains the training data for OneThinker, an all-in-one reasoning model for image and video, as presented in the paper OneThinker: All-in-one Reasoning Model for Image and Video.
Code: https://github.com/tulerfeng/OneThinker
About the OneThinker Dataset
OneThinker-600k is a large-scale multi-task training corpus designed to train OneThinker, an all-in-one multimodal reasoning model capable of understanding… See the full description on the dataset page: https://huggingface.co/datasets/OneThink/OneThinker-train-data.OneThinker-train-data
OneThinker-600k Training Data
This repository contains the training data for OneThinker, an all-in-one reasoning model for image and video, as presented in the paper OneThinker: All-in-one Reasoning Model for Image and Video.
Code: https://github.com/tulerfeng/OneThinker
About the OneThinker Dataset
OneThinker-600k is a large-scale multi-task training corpus designed to train OneThinker, an all-in-one multimodal reasoning model capable of understanding… See the full description on the dataset page: https://huggingface.co/datasets/luckywin90/OneThinker-train-data.One-to-All-sub
One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer
This repository contains the sample training data and benchmarks associated with the paper One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer.
The paper presents a unified framework for high-fidelity character animation and image pose transfer for references with arbitrary layouts, addressing spatial misalignment and partially visible references through innovative… See the full description on the dataset page: https://huggingface.co/datasets/MochunniaN1/One-to-All-sub.onetwovla-dataset
Datasets for OneTwoVLA
[Project Page] | [Paper] | [Code]
This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning.
The robot data covers two main tasks:
Cocktail
Open-World Visual Grounding
Dataset Folders
cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page: https://huggingface.co/datasets/Richard-Nai/onetwovla-dataset.OneThinker-evalThis repository contains the evaluation data presented in: OneThinker: All-in-one Reasoning Model for Image and Video
Code: https://github.com/tulerfeng/OneThinker
About OneThinker
We introduce OneThinker, an all-in-one multimodal reasoning generalist that is capable of thinking across a wide range of fundamental visual tasks within a single model.
We construct the large-scale OneThinker-600k multi-task training corpus and build OneThinker-SFT-340k with high-quality CoT… See the full description on the dataset page: https://huggingface.co/datasets/OneThink/OneThinker-eval.onetwovla
Datasets for OneTwoVLA
[Project Page] | [Paper] | [Code]
This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning.
The robot data covers two main tasks:
Cocktail
Open-World Visual Grounding
Dataset Folders
cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page: https://huggingface.co/datasets/yilin-wu/onetwovla.one-traj-demos-recollect-1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 80,
"total_frames": 15979,
"total_tasks": 67,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:80"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/one-traj-demos-recollect-1.LPFF
LPFF: Large-Pose-Flickr-Faces Dataset
LPFF is a large-pose Flickr face dataset comprised of 19,590 high-quality real large-pose portrait images.
[ICCV 2023] LPFF: A Portrait Dataset for Face Generators Across Large Poses
Yiqian Wu, Jing Zhang, Hongbo Fu, Xiaogang Jin*
Paper Video Suppl Project Page
The creation of 2D realistic facial images and 3D face shapes using generative networks has been a hot topic in recent years. Existing face… See the full description on the dataset page: https://huggingface.co/datasets/onethousand/LPFF.OneThinker_train_data_sftone-traj-demos-uploadThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 73,
"total_frames": 14862,
"total_tasks": 61,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:73"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/one-traj-demos-upload.one-traj-demos-prioritizedThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 90,
"total_frames": 18054,
"total_tasks": 74,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:90"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/one-traj-demos-prioritized.360degree-PHQ
[Preprint] 3DPortraitGAN: Learning One-Quarter Headshot 3D GANs from a Single-View Portrait Dataset with Diverse Body Poses
Yiqian Wu, Hao Xu, Xiangjun Tang, Hongbo Fu, Xiaogang Jin*
Paper (Arxiv) Supplementary (Google Drive)
This is the training dataset, 360°PHQ dataset, of 3DPortraitGAN.
one_twenty_five_faces
label_names = {
0: "Adriana Lima",
1: "Akshay Kumar",
2: "Alex Lawther",
3: "Alexandra Daddario",
4: "Alia Bhatt",
5: "Allen Page",
6: "Alvaro Morte",
7: "Alycia Debnam-Carey",
8: "Amanda Crew",
9: "Amber Heard",
10: "Amitabh Bachchan",
11: "Andy Samberg",
12: "Anne Hathaway",
13: "Anthony Mackie",
14: "Anushka Sharma",
15: "Avril Lavigne",
16: "Barack Obama",
17: "Barbara Palvin",
18: "Ben Affleck",
19: "Bill… See the full description on the dataset page: https://huggingface.co/datasets/jtatman/one_twenty_five_faces.Portrait3D_gallery
[SIGGRAPH 2024] Portrait3D: Text-Guided High-Quality 3D Portrait Generation Using Pyramid Representation and GANs Prior
Yiqian Wu, Hao Xu, Xiangjun Tang, Xien Chen, Siyu Tang, Zhebin Zhang, Chen Li, Xiaogang Jin*
We provide 440 3D portrait results along with their corresponding prompts, which are generated by our Portrait3D.
Data structure:
Portrait3D_gallery
│
└─── 000
│ │
│ └─── 000_pyramid_trigrid.pth [the pyramid trigrid file]
│ │
│ └─── 000_prompt.txt [the prompt]
│ │
│… See the full description on the dataset page: https://huggingface.co/datasets/onethousand/Portrait3D_gallery.one_token_at_a_time
One Token at a Time — companion data
Companion dataset for katha-ai/one_token_at_a_time,
the code release for Attending to Multimodal Generation One Token at a Time. This holds the task
CSVs, images, precomputed POS-tag outputs, and derived evaluation results needed to run and
reproduce the repo's five supported tasks: Fruit-Math, Fruit-Sport, Math-Fruit, VSR, and ChartQA.
Directory structure
data/
├── blank_black_image.png (not included — see note below)
├──… See the full description on the dataset page: https://huggingface.co/datasets/katha-ai-iiith/one_token_at_a_time.onetrainerdatapodonethinker_enfrom OneThink/OneThinker-train-data
AnimPortrait3D_gallery
AnimPortrait3D Results Gallery
This gallery showcases the results of AnimPortrait3D.
🔹 For interactive visualization, visit our GitHub page.
Preview Images
Preview images are available in the preview folder within this project.
File Structure
Each result is stored in a ZIP archive named after the face ID, containing the following files:
face_id.zip
│
├── fitted_params.pkl # Fitted SMPL-X parameters
│
├── point_cloud.ply # The generated avatar… See the full description on the dataset page: https://huggingface.co/datasets/onethousand/AnimPortrait3D_gallery.onetwovla-dataset
Datasets for OneTwoVLA
[Project Page] | [Paper] | [Code]
This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning.
The robot data covers two main tasks:
Cocktail
Open-World Visual Grounding
Dataset Folders
cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page: https://huggingface.co/datasets/ad1t7a/onetwovla-dataset.onethinker_koronethinker_kotranslate OneThink/OneThinker-train-data
OneThinker_img_train
