datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
movie_gen_video_bench
Dataset Card for the Movie Gen Benchmark
Movie Gen is a cast of foundation models that generates high-quality, 1080p HD videos with different aspect ratios and synchronized audio.
Here, we introduce our evaluation benchmark "Movie Gen Bench Video Bench", as detailed in the Movie Gen technical report (Section 3.5.2).
To enable fair and easy comparison to Movie Gen for future works on these evaluation benchmarks, we additionally release the non cherry-picked generated videos from… See the full description on the dataset page: https://huggingface.co/datasets/meta-ai-for-media-research/movie_gen_video_bench.Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated
Dataset Card for 5CD-AI/Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated
This translated dataset includes:
LLaVA-Video-178K: 178,509 caption entries, 960,791 open-ended QA (question and answer) items, and 196,198 multiple-choice QA items.
The video source of the original dataset is in this repo: lmms-lab/LLaVA-Video-178K
tgk-ai-video-generators-2026
Permanent dataset archive: https://doi.org/10.5281/zenodo.22703594
Five AI Video Generators Tested on Dialogue, Action and an Advert
These Guys Know tested Seedance 2.5, MiniMax H3, FLUX 3 Video, Gemini Omni 1.1 Flash and HappyHorse 1.1 on 1 September 2026. Every model received the same three ten-second, 16:9 text-to-video tasks: a father interrupting a computer game, a three-person fight inside a fixed hotel lobby and a Mango Cola advert with an exact product name.
We retained… See the full description on the dataset page: https://huggingface.co/datasets/These-Guys-Know/tgk-ai-video-generators-2026.AI-Vault-Videosai-generated_videoAI-Vault-Videosmovie_gen_video_bench_no_generations
Dataset Summary
Please see the full dataset huggingface page
ai-translate-video
Video Localization Timing Dataset
Dataset Summary
This dataset is a self-authored synthetic benchmark for multilingual video localization workflows. It focuses on the timing pressure that appears when subtitle segments are translated across languages and then reviewed for dubbing fit, subtitle-window preservation, and lip-sync risk. The package is designed for repository-safe experimentation and documentation. It does not contain third-party video, third-party audio… See the full description on the dataset page: https://huggingface.co/datasets/hellohihiloy789/ai-translate-video.egocentric-activity-video-samples
Egocentric Activity Video Samples
This sample shows first-person activity video for reviewing task flow, camera perspective, and real-world action structure before scoping a larger delivery.
What This Shows
Egocentric footage of everyday task activity
Clip-level metadata for task and scene review
A view of capture quality, framing, and movement patterns
Dataset Specifications
Field
Value
Modality
Video
Domain
First-person activity… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/egocentric-activity-video-samples.LearningChat_ai_video_production
Hallym AI Video Production Practice 2025-2 Public Dataset
1. 데이터셋 개요
데이터셋명: 한림대학교 AI영상제작실습 2025-2 공개용 데이터셋
교과목명: AI영상제작실습
학기: 2025-2
생성 배경: 2025학년도 2학기 AI영상제작실습 수업에서 조별로 제작·제출한 AI 기반 영상 결과물을 공개용 데이터셋 형태로 정리한 것이다.
목적: 수업 기반 AI 영상 창작 결과물을 공개 아카이브 형태로 정리하고, 작품 단위 메타데이터를 함께 제공하기 위함이다.
2. 데이터셋 범위
총 작품 수: 20편
데이터 단위: 조별 제출 영상 1편 = metadata.csv 1행
포함 대상: 1조부터 20조까지 각 팀 폴더의 원본 MP4 1개
제외 대상:
보고서 파일(.pdf, .docx, .hwp)
라이선스 동의서 파일
AI제작콘텐츠 발표회 2025 출품작 폴더에 따로 복사된 중복 MP4… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/LearningChat_ai_video_production.ai-video-cinematography-promptsvideo-generator-ai-agent
Video Generator Agent Meta and Traffic Dataset in AI Agent Marketplace | AI Agent Directory | AI Agent Index from DeepNLP
This dataset is collected from AI Agent Marketplace Index and Directory at http://www.deepnlp.org, which contains AI Agents's meta information such as agent's name, website, description, as well as the monthly updated Web performance metrics, including Google,Bing average search ranking positions, Github Stars, Arxiv References, etc.
The dataset is helpful for AI… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/video-generator-ai-agent.10000-Hour-Egocentric-Video-Dataset
10000-Hour-Egocentric-Video-Dataset
Description
This dataset contains 10,000 hours of egocentric multimodal data collected from diverse real-world environments, including residential, retail, and office scenarios. It covers a wide range of human activities and manipulation tasks, such as meal preparation, cleaning, storage, garment care, merchandising, and object picking. Each sample includes synchronized 4K stereo video, camera calibration parameters, 76-point… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-AI/10000-Hour-Egocentric-Video-Dataset.legal-videos-rag
Legal Videos RAG Benchmark
Legal Videos is a benchmark for evaluating RAG pipelines on real-world legal videos pulled from two legal proceedings video datasets.
LocalView, the largest known database of local government public meetings as they are captured and uploaded online covering more than 1000 hours of video.
Seattle City meetings from the Council Data Project (CDP), is the Seattle city subset of the CDP data having meeting videos and multiple metadata covering about 1200… See the full description on the dataset page: https://huggingface.co/datasets/aintropy-ai/legal-videos-rag.ai_tkt_video_testingShoji_ai_videos_captioned
first-person-task-video-samples
First-Person Task Video Samples
This sample set shows first-person task videos across household, outdoor, kitchen, light maintenance, and fine-motor activities. The clips are selected to make the task, hand-object interaction, and capture style easy to review before scoping broader coverage.
What This Shows
Egocentric video from practical daily-task scenarios
Clear hand-object interaction across five task categories
Compact sample clips selected for quick review… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/first-person-task-video-samples.behavioral-multimodal-video-samples
Behavioral Multimodal Video Samples
This sample shows multimodal human activity episodes for reviewing action coverage, visual context, and metadata structure before scoping a larger delivery.
What This Shows
Human activity video with task-level context
Metadata suited for reviewing action labels and capture structure
Signals for evaluating scene diversity and multimodal alignment
Dataset Specifications
Field
Value
Modality
Video… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/behavioral-multimodal-video-samples.Viet-Tiktok-VideoQA
Example by GIF 1:
Long description:
Video mở đầu với hình ảnh lá cờ đỏ sao vàng tung bay trên cột cờ cao, phía dưới là một quảng trường rộng lớn với nhiều người đi lại. Cảnh quay tiếp tục chuyển sang một buổi hoàng hôn rực rỡ với mặt trời đỏ cam đang lặn trên mặt nước, cùng lúc đó có nhiều người đang chèo thuyền kayak trên sông. Sau đó, ống kính hướng đến một tòa tháp cổ kính nhiều tầng ẩn mình giữa những tán cây xanh. Tiếp theo là hình ảnh một chàng trai trẻ đang đi bộ trên một… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-Tiktok-VideoQA.rav4-video-retrieval-index-AI-HW-2ai-video-leaderboard-2026
Which AI Video Generator Do AI Engines Recommend Most? (2026)
First-party research dataset from the Lattice network, published open under CC-BY 4.0 with a
permanent DOI. Nothing here is scraped from another dataset — it is computed and published under a
single ORCID-verified byline.
DOI
10.5281/zenodo.21272386
Published by
Nesyona (nesyona.com)
Study page
https://nesyona.com/research/ai-video-leaderboard-2026/
Licence
CC-BY 4.0 — reuse freely with attribution… See the full description on the dataset page: https://huggingface.co/datasets/vincentcouey/ai-video-leaderboard-2026.Suno-Public-Playlist-small-videoWe are sharing the videos of the subset of the audio we shared in our other dataset.
video_metadata_for_ai_trainingai-video-generation-benchmarks
