CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OneThink /OneThinker-train-data OneThinker-600k Training Data This repository contains the training data for OneThinker, an all-in-one reasoning model for image and video, as presented in the paper OneThinker: All-in-one Reasoning Model for Image and Video. Code: https://github.com/tulerfeng/OneThinker About the OneThinker Dataset OneThinker-600k is a large-scale multi-task training corpus designed to train OneThinker, an all-in-one multimodal reasoning model capable of understanding… See the full description on the dataset page: https://huggingface.co/datasets/OneThink/OneThinker-train-data.imageimage-text-to-text10K<n<100K14 likes5.2k downloads8mo agoHugging Face02luckywin90 /OneThinker-train-data OneThinker-600k Training Data This repository contains the training data for OneThinker, an all-in-one reasoning model for image and video, as presented in the paper OneThinker: All-in-one Reasoning Model for Image and Video. Code: https://github.com/tulerfeng/OneThinker About the OneThinker Dataset OneThinker-600k is a large-scale multi-task training corpus designed to train OneThinker, an all-in-one multimodal reasoning model capable of understanding… See the full description on the dataset page: https://huggingface.co/datasets/luckywin90/OneThinker-train-data.imageimage-text-to-text10K<n<100K1 likes2.5k downloads10mo agoHugging Face03MochunniaN1 /One-to-All-sub One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer This repository contains the sample training data and benchmarks associated with the paper One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer. The paper presents a unified framework for high-fidelity character animation and image pose transfer for references with arbitrary layouts, addressing spatial misalignment and partially visible references through innovative… See the full description on the dataset page: https://huggingface.co/datasets/MochunniaN1/One-to-All-sub.imageimage-to-video1K<n<10K4 likes914 downloads10mo agoHugging Face04Richard-Nai /onetwovla-dataset Datasets for OneTwoVLA [Project Page] | [Paper] | [Code] This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning. The robot data covers two main tasks: Cocktail Open-World Visual Grounding Dataset Folders cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page: https://huggingface.co/datasets/Richard-Nai/onetwovla-dataset.imagerobotics100K<n<1M7 likes709 downloads1y agoHugging Face05OneThink /OneThinker-evalThis repository contains the evaluation data presented in: OneThinker: All-in-one Reasoning Model for Image and Video Code: https://github.com/tulerfeng/OneThinker About OneThinker We introduce OneThinker, an all-in-one multimodal reasoning generalist that is capable of thinking across a wide range of fundamental visual tasks within a single model. We construct the large-scale OneThinker-600k multi-task training corpus and build OneThinker-SFT-340k with high-quality CoT… See the full description on the dataset page: https://huggingface.co/datasets/OneThink/OneThinker-eval.imageimage-text-to-text1 likes586 downloads10mo agoHugging Face06yilin-wu /onetwovla Datasets for OneTwoVLA [Project Page] | [Paper] | [Code] This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning. The robot data covers two main tasks: Cocktail Open-World Visual Grounding Dataset Folders cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page: https://huggingface.co/datasets/yilin-wu/onetwovla.imagerobotics100K<n<1M0 likes219 downloads1y agoHugging Face07windfromthenorth /one-traj-demos-recollect-1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "panda", "total_episodes": 80, "total_frames": 15979, "total_tasks": 67, "total_videos": 0, "total_chunks": 1, "chunks_size": 1000, "fps": 10, "splits": { "train": "0:80" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/one-traj-demos-recollect-1.imagerobotics10K<n<100K0 likes219 downloads7mo agoHugging Face08onethousand /LPFF LPFF: Large-Pose-Flickr-Faces Dataset LPFF is a large-pose Flickr face dataset comprised of 19,590 high-quality real large-pose portrait images. [ICCV 2023] LPFF: A Portrait Dataset for Face Generators Across Large Poses Yiqian Wu, Jing Zhang, Hongbo Fu, Xiaogang Jin* Paper Video Suppl Project Page The creation of 2D realistic facial images and 3D face shapes using generative networks has been a hot topic in recent years. Existing face… See the full description on the dataset page: https://huggingface.co/datasets/onethousand/LPFF.image4 likes206 downloads3y agoHugging Face09PerRing /OneThinker_train_data_sftimage100K<n<1M0 likes158 downloads8mo agoHugging Face10windfromthenorth /one-traj-demos-uploadThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "panda", "total_episodes": 73, "total_frames": 14862, "total_tasks": 61, "total_videos": 0, "total_chunks": 1, "chunks_size": 1000, "fps": 10, "splits": { "train": "0:73" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/one-traj-demos-upload.imagerobotics10K<n<100K0 likes119 downloads7mo agoHugging Face11windfromthenorth /one-traj-demos-prioritizedThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "panda", "total_episodes": 90, "total_frames": 18054, "total_tasks": 74, "total_videos": 0, "total_chunks": 1, "chunks_size": 1000, "fps": 10, "splits": { "train": "0:90" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/one-traj-demos-prioritized.imagerobotics10K<n<100K0 likes113 downloads7mo agoHugging Face12onethousand /360degree-PHQ [Preprint] 3DPortraitGAN: Learning One-Quarter Headshot 3D GANs from a Single-View Portrait Dataset with Diverse Body Poses Yiqian Wu, Hao Xu, Xiangjun Tang, Hongbo Fu, Xiaogang Jin* Paper (Arxiv) Supplementary (Google Drive) This is the training dataset, 360°PHQ dataset, of 3DPortraitGAN. image0 likes77 downloads2y agoHugging Face13jtatman /one_twenty_five_faces label_names = { 0: "Adriana Lima", 1: "Akshay Kumar", 2: "Alex Lawther", 3: "Alexandra Daddario", 4: "Alia Bhatt", 5: "Allen Page", 6: "Alvaro Morte", 7: "Alycia Debnam-Carey", 8: "Amanda Crew", 9: "Amber Heard", 10: "Amitabh Bachchan", 11: "Andy Samberg", 12: "Anne Hathaway", 13: "Anthony Mackie", 14: "Anushka Sharma", 15: "Avril Lavigne", 16: "Barack Obama", 17: "Barbara Palvin", 18: "Ben Affleck", 19: "Bill… See the full description on the dataset page: https://huggingface.co/datasets/jtatman/one_twenty_five_faces.image10K<n<100K0 likes53 downloads1y agoHugging Face14onethousand /Portrait3D_gallery [SIGGRAPH 2024] Portrait3D: Text-Guided High-Quality 3D Portrait Generation Using Pyramid Representation and GANs Prior Yiqian Wu, Hao Xu, Xiangjun Tang, Xien Chen, Siyu Tang, Zhebin Zhang, Chen Li, Xiaogang Jin* We provide 440 3D portrait results along with their corresponding prompts, which are generated by our Portrait3D. Data structure: Portrait3D_gallery │ └─── 000 │ │ │ └─── 000_pyramid_trigrid.pth [the pyramid trigrid file] │ │ │ └─── 000_prompt.txt [the prompt] │ │ │… See the full description on the dataset page: https://huggingface.co/datasets/onethousand/Portrait3D_gallery.imagen<1K7 likes47 downloads2y agoHugging Face15katha-ai-iiith /one_token_at_a_time One Token at a Time — companion data Companion dataset for katha-ai/one_token_at_a_time, the code release for Attending to Multimodal Generation One Token at a Time. This holds the task CSVs, images, precomputed POS-tag outputs, and derived evaluation results needed to run and reproduce the repo's five supported tasks: Fruit-Math, Fruit-Sport, Math-Fruit, VSR, and ChartQA. Directory structure data/ ├── blank_black_image.png (not included — see note below) ├──… See the full description on the dataset page: https://huggingface.co/datasets/katha-ai-iiith/one_token_at_a_time.imagevisual-question-answering1 likes43 downloads3d agoHugging Face16Meyem /onetrainerdatapodimagen<1K0 likes34 downloads1y agoHugging Face17KORMo-VL /onethinker_enfrom OneThink/OneThinker-train-data image100K<n<1M0 likes24 downloads7mo agoHugging Face18onethousand /AnimPortrait3D_gallery AnimPortrait3D Results Gallery This gallery showcases the results of AnimPortrait3D. 🔹 For interactive visualization, visit our GitHub page. Preview Images Preview images are available in the preview folder within this project. File Structure Each result is stored in a ZIP archive named after the face ID, containing the following files: face_id.zip │ ├── fitted_params.pkl # Fitted SMPL-X parameters │ ├── point_cloud.ply # The generated avatar… See the full description on the dataset page: https://huggingface.co/datasets/onethousand/AnimPortrait3D_gallery.3dtext-to-3dn<1K1 likes19 downloads1y agoHugging Face19ad1t7a /onetwovla-dataset Datasets for OneTwoVLA [Project Page] | [Paper] | [Code] This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning. The robot data covers two main tasks: Cocktail Open-World Visual Grounding Dataset Folders cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page: https://huggingface.co/datasets/ad1t7a/onetwovla-dataset.imagerobotics100K<n<1M0 likes17 downloads9mo agoHugging Face20KORMo-VLM /onethinker_korgatedimage10K<n<100K0 likes5 downloads8mo agoHugging Face21KORMo-VL /onethinker_kotranslate OneThink/OneThinker-train-data image100K<n<1M0 likes5 downloads7mo agoHugging Face22ShinDJ /OneThinker_img_traingatedimage100K<n<1M0 likes2 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.