CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01leduytho /vitra-ego4d-videovideo1K<n<10K4 likes6.5k downloads3mo agoHugging Face02Shiym /ViT-FineTuneimageimage-classification10K<n<100K0 likes5.2k downloads2y agoHugging Face03xincan /Llama-VITS_data Dataset Card for Llama-VITS_data The dataset repository contains data related with our work "Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness", encapsulating: Filtered dataset EmoV_DB_bea_sem Filelists with semantic embeddings Model checkpoints Human evaluation templates Dataset Details Paper: Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness Curated by: Xincan Feng, Akifumi Yoshimoto Funded by: CyberAgent Inc Repository:… See the full description on the dataset page: https://huggingface.co/datasets/xincan/Llama-VITS_data.text-to-speech2 likes4.7k downloads2y agoHugging Face04tals /vitaminc Details Fact Verification dataset created for Get Your Vitamin C! Robust Fact Verification with Contrastive Evidence (Schuster et al., NAACL 21`) based on Wikipedia edits (revisions). For more details see: https://github.com/TalSchuster/VitaminC When using this dataset, please cite the paper: BibTeX entry and citation info @inproceedings{schuster-etal-2021-get, title = "Get Your Vitamin {C}! Robust Fact Verification with Contrastive Evidence", author =… See the full description on the dataset page: https://huggingface.co/datasets/tals/vitaminc.texttext-classification100K<n<1M11 likes4k downloads4y agoHugging Face05Emanresu /features-dinov3-vith16plus-224-imagenet-22k-wdstext1M<n<10M0 likes2.3k downloads11mo agoHugging Face06VITRA-VLA /VITRA-1M VITRA-1M: Human Hand V-L-A Dataset Dataset Summary VITRA-1M is a large-scale Human Hand Visual-Language-Action (V-L-A) dataset constructed as described in the paper Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos. It contains 1.2 million short episodes with segmented language annotations, camera parameters (corrected intrinsics/extrinsics), and 3D hand reconstructions (left and right… See the full description on the dataset page: https://huggingface.co/datasets/VITRA-VLA/VITRA-1M.1M<n<10M29 likes2.2k downloads10mo agoHugging Face07MTDoven /ViTTiny1022 ViTTiny1022 The dataset for Scaling Up Parameter Generation: A Recurrent Diffusion Approach. Requirement Install torch and other dependencies conda install pytorch==2.3.1 torchvision==0.18.1 torchaudio==2.3.1 pytorch-cuda=12.1 -c pytorch -c nvidia pip install timm einops seaborn openpyxl Usage Test one checkpoint cd ViTTiny1022 python test.py ./chechpoint_test/0000_acc0.9613_class0314_condition_cifar10_vittiny.pth # python test.py… See the full description on the dataset page: https://huggingface.co/datasets/MTDoven/ViTTiny1022.1K<n<10K2 likes2.1k downloads2y agoHugging Face08quastAI /behavior-1k-2025-challenge-vjepa2-vitg-demo-embeddings V-JEPA 2 ViT-G Embeddings — BEHAVIOR-1K 2025 Challenge Demos (62h) Precomputed video embeddings for a 62-hour subsample of the BEHAVIOR-1K 2025 challenge demonstrations, extracted with the V-JEPA 2 ViT-g encoder. The goal is to make downstream experimentation faster and more reproducible by eliminating repeated video decoding and encoder forward passes — lowering the barrier for teams without access to large GPU clusters. Field Value Source dataset… See the full description on the dataset page: https://huggingface.co/datasets/quastAI/behavior-1k-2025-challenge-vjepa2-vitg-demo-embeddings.videofeature-extraction1M<n<10M2 likes1.2k downloads4mo agoHugging Face09lerobot-raw /stanford_mask_vit_raw0 likes1.2k downloads2y agoHugging Face10Joshua69 /vitac5000 likes1.1k downloads12d agoHugging Face11laion /CLIP-ViT-H-14-laion2B-s32B-b79K-all-checkpointsThis repository contains the intermediate checkpoints for the model https://huggingface.co/laion/CLIP-ViT-H-14-laion2B-s32B-b79K. Each "epoch" corresponds to an additional (32B / 256) samples seen, consituting total of 256 "epochs" The purpose of releasing these checkpoints and optimizer states is to enable analysis. For the first 121 "epochs", training was done with float16 mixed precision before switching to bfloat16 after a loss blow up. 1 likes1.1k downloads8mo agoHugging Face12open-index /vitco ViTco 165,847,195 Vietnamese documents from 4 public corpora, 370.2 GB of Parquet, one schema This dataset is the pinned public Vietnamese corpora as gao read them, every source put to one contract and one schema, before any cleaning. Contents What is it What is in it Where the text came from How it is laid out Reading it What you can build with it One row The columns What this repo is What ships and what does not Things to know before you use it What this is… See the full description on the dataset page: https://huggingface.co/datasets/open-index/vitco.text-generation100M<n<1B1 likes1k downloads1mo agoHugging Face13sjmathy /vitra-dinotxt-features0 likes981 downloads2mo agoHugging Face14Koushim /food101-vit-processedimage10K<n<100K0 likes823 downloads1y agoHugging Face15epfl-vita /svi-benchmark Stable Video Infinity (SVI) Benchmark Dataset This benchmark dataset is introduced in the paper: Stable Video Infinity: Infinite-Length Video Generation with Error Recycling by Wuyang Li, Wentao Pan, Po-Chien Luan, Yang Gao, Alexandre Alahi (2025). Project page: https://stable-video-infinity.github.io/homepage/ Code: https://github.com/vita-epfl/Stable-Video-Infinity Abstract We propose Stable Video Infinity (SVI) that is able to generate infinite-length videos with… See the full description on the dataset page: https://huggingface.co/datasets/epfl-vita/svi-benchmark.imageimage-to-videon<1K7 likes772 downloads11mo agoHugging Face16TryOnVirtual /VITON-HD-TESTimage1K<n<10K1 likes652 downloads3y agoHugging Face17closji /cc12m_openai_clip-vit-base-patch32_image_image_retrieval_pairs_2022-09-13image10M<n<100M0 likes640 downloads4y agoHugging Face18yusufani /TrCaption-trclip-vitl14-e101 likes627 downloads4y agoHugging Face19microsoft /VITRA-TeleData VITRA Teleoperation Dataset Dataset Summary This dataset contains real-world robot teleoperation demonstrations collected using a 7-DoF robotic arm equipped with a dexterous hand and a head-mounted RGB camera. Each episode provides synchronized numerical state/action data and video recordings. The dataset is used for finetuning in the project VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos Project… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/VITRA-TeleData.robotics1K<n<10K3 likes592 downloads8mo agoHugging Face20Qdrant /wolt-food-clip-ViT-B-32-embeddings wolt-food-clip-ViT-B-32-embeddings Qdrant's Food Discovery demo relies on the dataset of food images from the Wolt app. Each point in the collection represents a dish with a single image. The image is represented as a vector of 512 float numbers. Generation process The embeddings generated with clip-ViT-B-32 model have been generated using the following code snippet: from PIL import Image from sentence_transformers import SentenceTransformer image_path =… See the full description on the dataset page: https://huggingface.co/datasets/Qdrant/wolt-food-clip-ViT-B-32-embeddings.imagefeature-extraction1M<n<10M8 likes557 downloads3y agoHugging Face21vitaliy-sharandin /energy-consumption-hourly-spaintabular10K<n<100K3 likes555 downloads3y agoHugging Face22Rootreck /so-vits-svc-4.0-ru-The_Witcher_3_Wild_HuntЭто тренировочные данные моделей голосов персонажей из "Ведьмак 3: Дикая охота" для so-vits-svc-4.1.1 audio10K<n<100K0 likes555 downloads2y agoHugging Face23lscpku /VITATECSgated Dataset Card for VITATECS Dataset Description Dataset Summary VITATECS is a diagnostic VIdeo-Text dAtaset for the evaluation of TEmporal Concept underStanding. [2023/11/27] We have updated a new version of VITATECS which is generated using ChatGPT. The previous version generated by OPT-175B can be found here. Languages English. Dataset Structure Usage aspect = 'Type' #… See the full description on the dataset page: https://huggingface.co/datasets/lscpku/VITATECS.text10K<n<100K5 likes551 downloads2y agoHugging Face24pixxu /ViTextRender-500K Vietnamese Text Render 500K Dataset A large-scale dataset containing 500K Vietnamese text rendering image-text pairs for training generative models to improve text rendering performance. Dataset Structure image: Rendered text image in PNG format text: Corresponding text content filename: Original filename Usage This dataset is designed for fine-tuning generative models to improve text rendering capabilities on Vietnamese language. from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/pixxu/ViTextRender-500K.texttext-to-image100K<n<1M0 likes537 downloads10mo agoHugging Face25dari-ai /vite-selfbench Vite Selfbench Vite Selfbench is a 27-task software-engineering benchmark for coding agents, packaged for the Harbor evaluation framework. Each task asks an agent to implement a change in a frozen revision of vitejs/vite, then checks the resulting patch with task-specific tests. Harbor provides isolated task environments and runs verification separately from the agent. This repository contains the raw evaluation only. It does not include model outputs, scores, costs, or… See the full description on the dataset page: https://huggingface.co/datasets/dari-ai/vite-selfbench.n<1K0 likes530 downloads2mo agoHugging Face26meituan-longcat /VitaBench 🌱VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks 📃 Paper • 🌐 Website • 🏆 Leaderboard • 🛠️ Code • 🤗 Dataset 🔔 News [2026-01] Qwen3-Max-Thinking reported our Vita-Bench to evaluate and demonstrate its tool use capabilities (the averge score of 4 domains)!We invite the community to adopt Vita-Bench as the definitive touchstone for tool use performance assessment, and we appreciate diverse utilization & interpretation of our benchmark… See the full description on the dataset page: https://huggingface.co/datasets/meituan-longcat/VitaBench.28 likes502 downloads8mo agoHugging Face27open-index /vitco-clean ViTco Clean 2,279,914 Vietnamese documents from 2 public corpora, 10.6 GB of Parquet, one schema This dataset is the same corpora after the cleaning line: normalized, measured, filtered to Vietnamese prose, deduplicated on identity, and with the personal identifiers covered. Contents What is it What is in it Where the text came from How it is laid out Reading it What you can build with it One row The columns What this repo is What ships and what does not Things… See the full description on the dataset page: https://huggingface.co/datasets/open-index/vitco-clean.text-generation1M<n<10M0 likes501 downloads1mo agoHugging Face28Bupt-Joy /VitaSet VitaSet: Vision-Tactile VQA Dataset Overview VitaSet is a vision-tactile Visual Question Answering dataset for physical property reasoning. The dataset combines RGB vision and tactile sensing for material property understanding, containing 5,145 human-verified QA pairs across three tasks: hardness classification, material property description, and surface roughness classification. Hardware: Franka Emika Panda robot + GelSight Mini tactile sensor Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Bupt-Joy/VitaSet.imagevisual-question-answering10K<n<100K3 likes475 downloads9mo agoHugging Face29VItaldob /viciebski-pralietaryj-yiddish Viciebski Pralietaryj — Yiddish blocks Blocks of newspaper text set in Yiddish (Hebrew script), cut from scans of Viciebski pralietaryj («Віцебскі пралетарый»), a newspaper published in Vitebsk, Byelorussian SSR (Belarus), in 1930, 1931 and 1933. 2684 images from 78 pages across 78 issues. No transcriptions — this is a raw corpus for OCR/HTR work, not a labelled set. Belarusian blocks are present. The Yiddish pages ran as inserts inside a Belarusian newspaper, and blocks were… See the full description on the dataset page: https://huggingface.co/datasets/VItaldob/viciebski-pralietaryj-yiddish.imageimage-to-text1K<n<10K0 likes468 downloads15d agoHugging Face30NXN-Labs /VITON-HD-edit VITON-HD-edit This repository contains the VITON-HD-edit dataset presented in the paper CtrlVTON: Controllable Virtual Try-On via Visual-Instance-Prompt Segmentation. Github Repository | Paper (arXiv) Dataset Overview The VITON-HD-edit dataset is a public benchmark built to support three evaluations: Image-editing Virtual Try-On (VTO) Instance-level visual-prompt segmentation Spatial controllability Training the editing model requires triplets $(p… See the full description on the dataset page: https://huggingface.co/datasets/NXN-Labs/VITON-HD-edit.imageimage-to-image1K<n<10K0 likes463 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.