datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VideoArtifactDetectionVideo Artifact Detection dataset using source videos from LongVideoBench. Video level labels are given in labels.csv (training) and labels_test.csv (testing). Localized artifact regions for burst artifacts are given in the artifact_ranges column.
Note that labels.csv contains additional source videos from LongVideoBench that are not included in this repository. You may visit the LongVideoBench page for the additional videos.
For additional questions, please email palmerla@usc.edu.
VideoDRyodas_fa_nosub_videoidsVideoGen-RewardBench
🏆 [VideoGen-RewardBench Leaderboard]
Introduction
VideoGen-RewardBench is a comprehensive benchmark designed to evaluate the performance of video reward models on modern text-to-video (T2V) systems. Derived from the third-party VideoGen-Eval (Zeng et.al, 2024), we constructing 26.5k (prompt, Video A, Video B) triplets and employing expert annotators to provide pairwise preference labels.
These annotations are based on key evaluation dimensions—Visual Quality (VQ), Motion… See the full description on the dataset page: https://huggingface.co/datasets/KlingTeam/VideoGen-RewardBench.videophy2_testProject: https://github.com/Hritikbansal/videophy/tree/main/VIDEOPHY2
caption: original prompt in the dataset
video_url: generated video (using original prompt or upsampled caption, depending on the video model)
sa: semantic adherence score (1-5) from human evaluation
pc: physical commonsense score (1-5) from human evaluation
joint: computed as sa >= 4, pc >= 4
physics_rules_followed: list of physics rules followed in the video as judged by human annotators (1)
physics_rules_unfollowed: list… See the full description on the dataset page: https://huggingface.co/datasets/videophysics/videophy2_test.amazon_product_reviews_video_games#Title
videophy_test_publicWe have uploaded the videos at: https://huggingface.co/videophysics/videophy-test-videos/tree/main
For more details, please visit:
project github: https://github.com/Hritikbansal/videophy
project website: https://videophy.github.io/
tartanair_videotgk-ai-video-generators-2026
Permanent dataset archive: https://doi.org/10.5281/zenodo.22703594
Five AI Video Generators Tested on Dialogue, Action and an Advert
These Guys Know tested Seedance 2.5, MiniMax H3, FLUX 3 Video, Gemini Omni 1.1 Flash and HappyHorse 1.1 on 1 September 2026. Every model received the same three ten-second, 16:9 text-to-video tasks: a father interrupting a computer game, a three-person fight inside a fixed hotel lobby and a Mango Cola advert with an exact product name.
We retained… See the full description on the dataset page: https://huggingface.co/datasets/These-Guys-Know/tgk-ai-video-generators-2026.ViSignLanguage-Video
AI Challenge CV Dataset Description
This dataset comes from https://aichallenge.ptit.edu.vn/ and is organized for a multi-class video classification task focusing on Vietnamese Sign Language.
🛠 Dataset Viewer Configuration
This dataset supports Hugging Face Dataset Viewer. The dataset structure and video paths are mapped via dataset_metadata.csv.
Expected CSV Structure
To ensure the Dataset Viewer renders correctly, your dataset_metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/star092304/ViSignLanguage-Video.UCF101-Videosvideo_saliency
SalTempto — Saliency Eye-Tracking Dataset
A video eye-tracking dataset of 224 naturalistic video stimuli with gaze recordings from multiple subjects, designed for training and evaluating visual saliency models.
Dataset Overview
Videos
224 (1920×1080, ~30 fps)
Train / Val / Test
204 / 10 / 10
Subjects per video
Train: 1–3 (mean ≈ 2.5) · Val: 15–16 · Test: held out
Gaze recordings for the test split are intentionally held out as a hidden… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-saltempto-submission/video_saliency.video-data
TalorData Video Data
Rich, up-to-date video metadata + ready-to-transcribe MP4s from YouTube.
TalorData Video Data is a large, constantly refreshed dataset of YouTube videos. Each record bundles the video's pre-cut MP4 alongside an auto-generated, timestamped transcript — ready for fine-tuning, RAG, search, summarization, and multimodal inference.
Sample Files (Preview)
This repository hosts a public preview/sample of the record schema. Open sample-metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/TalorDataHQ/video-data.ViSignLanguage-Video
AI Challenge CV Dataset Description
This dataset comes from https://aichallenge.ptit.edu.vn/ and is organized for a multi-class video classification task focusing on Vietnamese Sign Language.
🛠 Dataset Viewer Configuration
This dataset supports Hugging Face Dataset Viewer. The dataset structure and video paths are mapped via dataset_metadata.csv.
Expected CSV Structure
To ensure the Dataset Viewer renders correctly, your dataset_metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/namwu/ViSignLanguage-Video.videophy2_upsampled_promptsTiktok-Videos
TikTok Video Analytics Dataset
Sample TikTok video dataset with comprehensive engagement metrics and metadata. Each row represents a single TikTok video with content and detailed analytics.
This is a sample dataset. To access the full version or request any custom dataset tailored to your needs, contact DataHive at contact@datahive.ai.
Files Included
train.csv – TikTok video analytics data
What's included
Video URLs and identifiers
Comprehensive engagement… See the full description on the dataset page: https://huggingface.co/datasets/datahiveai/Tiktok-Videos.VideoDPO-10k@misc{liu2024videodpoomnipreferencealignmentvideo,
title={VideoDPO: Omni-Preference Alignment for Video Diffusion Generation},
author={Runtao Liu and Haoyu Wu and Zheng Ziqiang and Chen Wei and Yingqing He and Renjie Pi and Qifeng Chen},
year={2024},
eprint={2412.14167},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2412.14167},
}
@misc{wang2024vidprommillionscalerealpromptgallery,
title={VidProM: A Million-scale Real… See the full description on the dataset page: https://huggingface.co/datasets/chungimungi/VideoDPO-10k.timeseries_trending_youtube_videos_2019-04-15_to_2020-04-15Timeseries Trending YouTube Videos: 2019-04-15 to 2020-04-15
This dataset is a csv of one of the archived historical database tables queried from my non public database that contains time series data for period of 2019-04-15 to 2020-04-15. Video data was captured from the time they first appeared on trending list, and TSD exists until the video is removed from trending list.
This snapshot contains data for the 11,369 videos that appeared on trending within the timeframe, with 1,541,128 records… See the full description on the dataset page: https://huggingface.co/datasets/jettisonthenet/timeseries_trending_youtube_videos_2019-04-15_to_2020-04-15.morse-500
MORSE-500 Benchmark
🔥 News
May 15, 2025: We release MORSE-500, 500 programmatically generated videos across six reasoning categories: abstract, mathematical, physical, planning, spatial, and temporal, to stress-test multimodal reasoning. Frontier models including OpenAI o3 and Gemini 2.5 Pro score lower than… See the full description on the dataset page: https://huggingface.co/datasets/video-reasoning/morse-500.VideoMind
🔍VideoMind: An Omni-Modal Video Dataset with Intent Grounding for Deep-Cognitive Video Understanding
Dataset Description
VideoMind is a large-scale video-centric multimodal dataset that can be used to learn powerful and transferable text-video representations
for video understanding tasks such as video question answering and video retrieval. The VideoMind dataset contains 105K(5K test for
only) video samples, each of which is accompanied by audio, as well as systematic… See the full description on the dataset page: https://huggingface.co/datasets/DixinChen/VideoMind.videocon_synVideo-Games-Sales-EDA
🎮 Video Game Sales — Exploratory Data Analysis (EDA)
Video Presentation
If I didn’t cover everything it’s because I didn’t have enough time
Your browser does not support the video tag.
Executive Summary
The video game industry is a multi‑billion dollar market characterized by extreme unpredictability—a single "mega‑hit" can generate more revenue than thousands of average games combined. This project analyzes historical video game sales… See the full description on the dataset page: https://huggingface.co/datasets/0tizm0/Video-Games-Sales-EDA.videogames_salesbest-selling-video-games
Dataset Card for [best-selling-video-games]
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/arjunpatel/best-selling-video-games.abhishekgupta56447_video-games-sales-from-zenodo
Video Games Sales (from Zenodo)
Video game sales data including platform, genre, publisher, and global sales.
Dataset Info
Source: Kaggle
Original Size: 0.37 MB
Kaggle Downloads: 819
Files: 1
Files
vgsales.csv
Mirrored from Kaggle
egocentric-video
Egocentric Dataset for Physical AI and Robotics
The dataset contains 4,050 hours of first-person videos for egocentric vision and egocentric tracking. Featuring multimodal data from egocentric views, it includes data annotations and motion capture for extracting 3d poses. It provides detailed 3d objects and 3d scenes using visual data from VR headsets to analyze hands motions and pose estimations.… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/egocentric-video.Nagi_no_Asukara_Videos_Captioned
Reorganized version of Wild-Heart/Disney-VideoGeneration-Dataset. This is needed for Mochi-1 fine-tuning.
class-to-video-prompt-generationVideoConviction
VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations
Paper: VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations
Conference: ACM SIGKDD 2025
Authors: Michael Galarnyk, Veer Kejriwal, Agam Shah, Yash Bhardwaj, Nicholas Watney Meyer, Anand Krishnan, Sudheer Chava
Dataset Summary
VideoConviction is a multimodal dataset curated from financial influencer (“finfluencer”) videos on YouTube. It… See the full description on the dataset page: https://huggingface.co/datasets/gtfintechlab/VideoConviction.seedance-2-5-free-tier-data
Seedance 2.5 free-tier observations
A dated record of what each channel actually grants on the Seedance 2.5 free tier — daily credits, daily generation caps, sign-up bonuses, clip length, extension ceiling, output resolution, watermark behaviour, and API availability.
Free-tier terms are the least documented part of an AI video product. Vendors publish a launch post with a headline number, then adjust the daily grant, move the watermark switch, or gate a model behind a… See the full description on the dataset page: https://huggingface.co/datasets/videofreetier/seedance-2-5-free-tier-data.
