CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01larrylarrylarry1 /VideoArtifactDetectionVideo Artifact Detection dataset using source videos from LongVideoBench. Video level labels are given in labels.csv (training) and labels_test.csv (testing). Localized artifact regions for burst artifacts are given in the artifact_ranges column. Note that labels.csv contains additional source videos from LongVideoBench that are not included in this repository. You may visit the LongVideoBench page for the additional videos. For additional questions, please email palmerla@usc.edu. tabular1K<n<10K0 likes2.4k downloads4mo agoHugging Face02strike20023 /VideoDRtextn<1K2 likes1.7k downloads9d agoHugging Face03saeedzouashkiani /yodas_fa_nosub_videoidstext10M<n<100M0 likes729 downloads1mo agoHugging Face04KlingTeam /VideoGen-RewardBench 🏆 [VideoGen-RewardBench Leaderboard] Introduction VideoGen-RewardBench is a comprehensive benchmark designed to evaluate the performance of video reward models on modern text-to-video (T2V) systems. Derived from the third-party VideoGen-Eval (Zeng et.al, 2024), we constructing 26.5k (prompt, Video A, Video B) triplets and employing expert annotators to provide pairwise preference labels. These annotations are based on key evaluation dimensions—Visual Quality (VQ), Motion… See the full description on the dataset page: https://huggingface.co/datasets/KlingTeam/VideoGen-RewardBench.tabular10K<n<100K9 likes514 downloads2y agoHugging Face05videophysics /videophy2_testProject: https://github.com/Hritikbansal/videophy/tree/main/VIDEOPHY2 caption: original prompt in the dataset video_url: generated video (using original prompt or upsampled caption, depending on the video model) sa: semantic adherence score (1-5) from human evaluation pc: physical commonsense score (1-5) from human evaluation joint: computed as sa >= 4, pc >= 4 physics_rules_followed: list of physics rules followed in the video as judged by human annotators (1) physics_rules_unfollowed: list… See the full description on the dataset page: https://huggingface.co/datasets/videophysics/videophy2_test.tabularvideo-classification1K<n<10K5 likes380 downloads8mo agoHugging Face06LoganKells /amazon_product_reviews_video_games#Title tabular10K<n<100K9 likes341 downloads5y agoHugging Face07videophysics /videophy_test_publicWe have uploaded the videos at: https://huggingface.co/videophysics/videophy-test-videos/tree/main For more details, please visit: project github: https://github.com/Hritikbansal/videophy project website: https://videophy.github.io/ tabular1K<n<10K1 likes280 downloads2y agoHugging Face08Alexhe101 /tartanair_videotext1K<n<10K0 likes274 downloads1y agoHugging Face09These-Guys-Know /tgk-ai-video-generators-2026 Permanent dataset archive: https://doi.org/10.5281/zenodo.22703594 Five AI Video Generators Tested on Dialogue, Action and an Advert These Guys Know tested Seedance 2.5, MiniMax H3, FLUX 3 Video, Gemini Omni 1.1 Flash and HappyHorse 1.1 on 1 September 2026. Every model received the same three ten-second, 16:9 text-to-video tasks: a father interrupting a computer game, a three-person fight inside a fixed hotel lobby and a Mango Cola advert with an exact product name. We retained… See the full description on the dataset page: https://huggingface.co/datasets/These-Guys-Know/tgk-ai-video-generators-2026.tabulartext-to-videon<1K0 likes193 downloads10d agoHugging Face10star092304 /ViSignLanguage-Video AI Challenge CV Dataset Description This dataset comes from https://aichallenge.ptit.edu.vn/ and is organized for a multi-class video classification task focusing on Vietnamese Sign Language. 🛠 Dataset Viewer Configuration This dataset supports Hugging Face Dataset Viewer. The dataset structure and video paths are mapped via dataset_metadata.csv. Expected CSV Structure To ensure the Dataset Viewer renders correctly, your dataset_metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/star092304/ViSignLanguage-Video.tabularvideo-classification1K<n<10K2 likes179 downloads4mo agoHugging Face11bitmind /UCF101-Videostext10K<n<100K0 likes178 downloads1y agoHugging Face12anonymous-saltempto-submission /video_saliency SalTempto — Saliency Eye-Tracking Dataset A video eye-tracking dataset of 224 naturalistic video stimuli with gaze recordings from multiple subjects, designed for training and evaluating visual saliency models. Dataset Overview Videos 224 (1920×1080, ~30 fps) Train / Val / Test 204 / 10 / 10 Subjects per video Train: 1–3 (mean ≈ 2.5) · Val: 15–16 · Test: held out Gaze recordings for the test split are intentionally held out as a hidden… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-saltempto-submission/video_saliency.textn<1K0 likes170 downloads5mo agoHugging Face13TalorDataHQ /video-data TalorData Video Data Rich, up-to-date video metadata + ready-to-transcribe MP4s from YouTube. TalorData Video Data is a large, constantly refreshed dataset of YouTube videos. Each record bundles the video's pre-cut MP4 alongside an auto-generated, timestamped transcript — ready for fine-tuning, RAG, search, summarization, and multimodal inference. Sample Files (Preview) This repository hosts a public preview/sample of the record schema. Open sample-metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/TalorDataHQ/video-data.imagevideo-classificationn<1K0 likes168 downloads22d agoHugging Face14namwu /ViSignLanguage-Video AI Challenge CV Dataset Description This dataset comes from https://aichallenge.ptit.edu.vn/ and is organized for a multi-class video classification task focusing on Vietnamese Sign Language. 🛠 Dataset Viewer Configuration This dataset supports Hugging Face Dataset Viewer. The dataset structure and video paths are mapped via dataset_metadata.csv. Expected CSV Structure To ensure the Dataset Viewer renders correctly, your dataset_metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/namwu/ViSignLanguage-Video.tabularvideo-classification1K<n<10K2 likes162 downloads3mo agoHugging Face15videophysics /videophy2_upsampled_promptstext1K<n<10K0 likes160 downloads2y agoHugging Face16datahiveai /Tiktok-Videos TikTok Video Analytics Dataset Sample TikTok video dataset with comprehensive engagement metrics and metadata. Each row represents a single TikTok video with content and detailed analytics. This is a sample dataset. To access the full version or request any custom dataset tailored to your needs, contact DataHive at contact@datahive.ai. Files Included train.csv – TikTok video analytics data What's included Video URLs and identifiers Comprehensive engagement… See the full description on the dataset page: https://huggingface.co/datasets/datahiveai/Tiktok-Videos.tabular1K<n<10K0 likes156 downloads1y agoHugging Face17chungimungi /VideoDPO-10k@misc{liu2024videodpoomnipreferencealignmentvideo, title={VideoDPO: Omni-Preference Alignment for Video Diffusion Generation}, author={Runtao Liu and Haoyu Wu and Zheng Ziqiang and Chen Wei and Yingqing He and Renjie Pi and Qifeng Chen}, year={2024}, eprint={2412.14167}, archivePrefix={arXiv}, primaryClass={cs.CV}, url={https://arxiv.org/abs/2412.14167}, } @misc{wang2024vidprommillionscalerealpromptgallery, title={VidProM: A Million-scale Real… See the full description on the dataset page: https://huggingface.co/datasets/chungimungi/VideoDPO-10k.textvideo-classification10K<n<100K1 likes154 downloads2y agoHugging Face18jettisonthenet /timeseries_trending_youtube_videos_2019-04-15_to_2020-04-15Timeseries Trending YouTube Videos: 2019-04-15 to 2020-04-15 This dataset is a csv of one of the archived historical database tables queried from my non public database that contains time series data for period of 2019-04-15 to 2020-04-15. Video data was captured from the time they first appeared on trending list, and TSD exists until the video is removed from trending list. This snapshot contains data for the 11,369 videos that appeared on trending within the timeframe, with 1,541,128 records… See the full description on the dataset page: https://huggingface.co/datasets/jettisonthenet/timeseries_trending_youtube_videos_2019-04-15_to_2020-04-15.tabular1M<n<10M6 likes148 downloads4y agoHugging Face19video-reasoning /morse-500 MORSE-500 Benchmark 🔥 News May 15, 2025: We release MORSE-500, 500 programmatically generated videos across six reasoning categories: abstract, mathematical, physical, planning, spatial, and temporal, to stress-test multimodal reasoning. Frontier models including OpenAI o3 and Gemini 2.5 Pro score lower than… See the full description on the dataset page: https://huggingface.co/datasets/video-reasoning/morse-500.textvideo-classificationn<1K2 likes137 downloads1y agoHugging Face20DixinChen /VideoMind 🔍VideoMind: An Omni-Modal Video Dataset with Intent Grounding for Deep-Cognitive Video Understanding Dataset Description VideoMind is a large-scale video-centric multimodal dataset that can be used to learn powerful and transferable text-video representations for video understanding tasks such as video question answering and video retrieval. The VideoMind dataset contains 105K(5K test for only) video samples, each of which is accompanied by audio, as well as systematic… See the full description on the dataset page: https://huggingface.co/datasets/DixinChen/VideoMind.textquestion-answering100K<n<1M1 likes135 downloads1y agoHugging Face21lucazanella /videocon_syntabular100K<n<1M0 likes135 downloads1y agoHugging Face220tizm0 /Video-Games-Sales-EDA 🎮 Video Game Sales — Exploratory Data Analysis (EDA) Video Presentation If I didn’t cover everything it’s because I didn’t have enough time Your browser does not support the video tag. Executive Summary The video game industry is a multi‑billion dollar market characterized by extreme unpredictability—a single "mega‑hit" can generate more revenue than thousands of average games combined. This project analyzes historical video game sales… See the full description on the dataset page: https://huggingface.co/datasets/0tizm0/Video-Games-Sales-EDA.tabulartabular-classification10K<n<100K0 likes126 downloads6mo agoHugging Face23maxfactor71 /videogames_salestabular10K<n<100K0 likes103 downloads5mo agoHugging Face24arjunpatel /best-selling-video-games Dataset Card for [best-selling-video-games] Dataset Summary [More Information Needed] Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation Curation Rationale [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/arjunpatel/best-selling-video-games.textn<1K1 likes94 downloads4y agoHugging Face25jason1966 /abhishekgupta56447_video-games-sales-from-zenodo Video Games Sales (from Zenodo) Video game sales data including platform, genre, publisher, and global sales. Dataset Info Source: Kaggle Original Size: 0.37 MB Kaggle Downloads: 819 Files: 1 Files vgsales.csv Mirrored from Kaggle tabular10K<n<100K0 likes91 downloads6mo agoHugging Face26UniDataPro /egocentric-video Egocentric Dataset for Physical AI and Robotics The dataset contains 4,050 hours of first-person videos for egocentric vision and egocentric tracking. Featuring multimodal data from egocentric views, it includes data annotations and motion capture for extracting 3d poses. It provides detailed 3d objects and 3d scenes using visual data from VR headsets to analyze hands motions and pose estimations.… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/egocentric-video.textroboticsn<1K1 likes85 downloads1mo agoHugging Face27svjack /Nagi_no_Asukara_Videos_Captioned Reorganized version of Wild-Heart/Disney-VideoGeneration-Dataset. This is needed for Mochi-1 fine-tuning. text1K<n<10K0 likes82 downloads2y agoHugging Face28shreyahegde /class-to-video-prompt-generationtexttext-generationn<1K1 likes76 downloads2y agoHugging Face29gtfintechlab /VideoConviction VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations Paper: VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations Conference: ACM SIGKDD 2025 Authors: Michael Galarnyk, Veer Kejriwal, Agam Shah, Yash Bhardwaj, Nicholas Watney Meyer, Anand Krishnan, Sudheer Chava Dataset Summary VideoConviction is a multimodal dataset curated from financial influencer (“finfluencer”) videos on YouTube. It… See the full description on the dataset page: https://huggingface.co/datasets/gtfintechlab/VideoConviction.tabularvideo-classificationn<1K3 likes76 downloads1y agoHugging Face30videofreetier /seedance-2-5-free-tier-data Seedance 2.5 free-tier observations A dated record of what each channel actually grants on the Seedance 2.5 free tier — daily credits, daily generation caps, sign-up bonuses, clip length, extension ceiling, output resolution, watermark behaviour, and API availability. Free-tier terms are the least documented part of an AI video product. Vendors publish a launch post with a headline number, then adjust the daily grant, move the watermark switch, or gate a model behind a… See the full description on the dataset page: https://huggingface.co/datasets/videofreetier/seedance-2-5-free-tier-data.textn<1K0 likes71 downloads6d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.