datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
physical-ai-bench-step-30000-videos
PhysicalAIBench Step 30000 Model Comparison
This repository contains two filename-aligned sets of 5,220 MP4 outputs from
step_30000 evaluation runs. The files are presented through the
companion PhysicalAI Video Gallery.
Model sets
Directory
Model
Files
Bytes
videos/
DC-AE v0.2, Cosmos encoder + causal decoder
5,220
4,244,965,541
videos_wan22_vae/
Wan 2.2 VAE, phase 3 PDX
5,220
4,818,900,749
The two directories have an exact 1:1 basename match.… See the full description on the dataset page: https://huggingface.co/datasets/spongy/physical-ai-bench-step-30000-videos.AudioSet_balanced_videomovie_gen_video_bench
Dataset Card for the Movie Gen Benchmark
Movie Gen is a cast of foundation models that generates high-quality, 1080p HD videos with different aspect ratios and synchronized audio.
Here, we introduce our evaluation benchmark "Movie Gen Bench Video Bench", as detailed in the Movie Gen technical report (Section 3.5.2).
To enable fair and easy comparison to Movie Gen for future works on these evaluation benchmarks, we additionally release the non cherry-picked generated videos from… See the full description on the dataset page: https://huggingface.co/datasets/meta-ai-for-media-research/movie_gen_video_bench.quantum-video
Dataset Card for Dataset Name
quantum suite video
quantum suite video dataset, used to train the quantum suite video model.
Dataset Details
will upload, collecting data
Dataset Description
This dataset is aimed to be curated for the quantum suite video dataset with video from wikimedia (Cc allowing commercial use), this dataset allows commercial use, and will become useful for you to use in your video models,
the size is aimed to be 2.5tb, after we… See the full description on the dataset page: https://huggingface.co/datasets/ai-api-key-free-finder/quantum-video.ai-shorts-videoswat.ai-so101-videosStep-Video-T2V-EvalThis dataset contains the data of the paper Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model.
Code: https://github.com/stepfun-ai/Step-Video-T2V
Project page: https://yuewen.cn/videos
Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated
Dataset Card for 5CD-AI/Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated
This translated dataset includes:
LLaVA-Video-178K: 178,509 caption entries, 960,791 open-ended QA (question and answer) items, and 196,198 multiple-choice QA items.
The video source of the original dataset is in this repo: lmms-lab/LLaVA-Video-178K
tgk-ai-video-generators-2026
Permanent dataset archive: https://doi.org/10.5281/zenodo.22703594
Five AI Video Generators Tested on Dialogue, Action and an Advert
These Guys Know tested Seedance 2.5, MiniMax H3, FLUX 3 Video, Gemini Omni 1.1 Flash and HappyHorse 1.1 on 1 September 2026. Every model received the same three ten-second, 16:9 text-to-video tasks: a father interrupting a computer game, a three-person fight inside a fixed hotel lobby and a Mango Cola advert with an exact product name.
We retained… See the full description on the dataset page: https://huggingface.co/datasets/These-Guys-Know/tgk-ai-video-generators-2026.videobridge-ai-kb
🎬 VideoBridge AI — Knowledge Base (RAG)
The single knowledge base for the VideoBridge AI project — used by the Telegram bot (@VideoBrige_bot) and the site panel for RAG search (/ai/rag).
This is the only Hugging Face dataset created for this project (no duplicate models/spaces/databases).
🔄 Auto-update
A daily scraper (scraper/kb_scrape.py) refreshes this dataset with:
📚 Academic papers (arXiv): text-to-video, video generation, video diffusion, RAG
📰 AI/tech… See the full description on the dataset page: https://huggingface.co/datasets/sosa123454321/videobridge-ai-kb.AI-Vault-VideosOld-AI-Video
📁 Old AI Video — Архив оригинальных AI-видео
🔒 Правообладатель
Абдулаев Самад Германович
📧 usnul.noxil@gmail.com
📋 Что это
Данный датасет является полным эталонным архивом всех оригинальных AI-видео и связанных материалов, созданных Абдулаевым Самадом Германовичем в период с 21 декабря 2024 года и опубликованных на Telegram-канале «ИИ-видео» (Telegram ID: -1002330425412).
Архив создан для:
✅ Фиксации авторства — доказательство того, что все материалы… See the full description on the dataset page: https://huggingface.co/datasets/HelioAI/Old-AI-Video.VIP-200K-Video
VIP-200K-Video
Overview
This repository contains only the video files (.tar archives) from the VIP-200K dataset.
For the complete dataset (including JSON annotations, face frames, and segment metadata), please download HiDream-ai/VIP-200K.
File Structure
video/
├── vip200k_train_0001_of_0100.tar
├── vip200k_train_0002_of_0100.tar
├── ...
└── vip200k_train_0100_of_0100.tar
Each .tar archive contains video clips organized by YouTube video ID:
{video_id}/… See the full description on the dataset page: https://huggingface.co/datasets/HiDream-ai/VIP-200K-Video.ai-generated_videoUGC-VideoCap
UGC-VideoCaptioner Dataset
Real-world user-generated videos, especially on platforms like TikTok, often feature rich and intertwined audio-visual content. However, existing video captioning benchmarks and models remain predominantly visual-centric, overlooking the crucial role of audio in conveying scene dynamics, speaker intent, and narrative context. This lack of full-modality datasets and lightweight, capable models hampers progress in fine-grained, multimodal video… See the full description on the dataset page: https://huggingface.co/datasets/Memories-ai/UGC-VideoCap.AI-Vault-Videosai_cctv_videosmovie_gen_video_bench_no_generations
Dataset Summary
Please see the full dataset huggingface page
ai-translate-video
Video Localization Timing Dataset
Dataset Summary
This dataset is a self-authored synthetic benchmark for multilingual video localization workflows. It focuses on the timing pressure that appears when subtitle segments are translated across languages and then reviewed for dubbing fit, subtitle-window preservation, and lip-sync risk. The package is designed for repository-safe experimentation and documentation. It does not contain third-party video, third-party audio… See the full description on the dataset page: https://huggingface.co/datasets/hellohihiloy789/ai-translate-video.AudioSet_unbalanced_videovideo_upscale_engineTelugu-NLP-AI-Dialect-Comedy-video-DatasetTelugu is one of the sweetest and oldest languages of India. A deep Dive into Telugu its spoken in 2 states and majorly 16 regional dailects.
This Dataset help you perform operations in NLP and Speech Recognition Models towards telugu Dialects.
Telugu-NLP-AI-Dialect-Comedy-video-DatasetTelugu is one of the sweetest and oldest languages of India. A deep Dive into Telugu its spoken in 2 states and majorly 16 regional dailects.
This Dataset help you perform operations in NLP and Speech Recognition Models towards telugu Dialects.
egocentric-activity-video-samples
Egocentric Activity Video Samples
This sample shows first-person activity video for reviewing task flow, camera perspective, and real-world action structure before scoping a larger delivery.
What This Shows
Egocentric footage of everyday task activity
Clip-level metadata for task and scene review
A view of capture quality, framing, and movement patterns
Dataset Specifications
Field
Value
Modality
Video
Domain
First-person activity… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/egocentric-activity-video-samples.ai-video-model-benchmarkai_videos_Financekaggle-ai-videoLearningChat_ai_video_production
Hallym AI Video Production Practice 2025-2 Public Dataset
1. 데이터셋 개요
데이터셋명: 한림대학교 AI영상제작실습 2025-2 공개용 데이터셋
교과목명: AI영상제작실습
학기: 2025-2
생성 배경: 2025학년도 2학기 AI영상제작실습 수업에서 조별로 제작·제출한 AI 기반 영상 결과물을 공개용 데이터셋 형태로 정리한 것이다.
목적: 수업 기반 AI 영상 창작 결과물을 공개 아카이브 형태로 정리하고, 작품 단위 메타데이터를 함께 제공하기 위함이다.
2. 데이터셋 범위
총 작품 수: 20편
데이터 단위: 조별 제출 영상 1편 = metadata.csv 1행
포함 대상: 1조부터 20조까지 각 팀 폴더의 원본 MP4 1개
제외 대상:
보고서 파일(.pdf, .docx, .hwp)
라이선스 동의서 파일
AI제작콘텐츠 발표회 2025 출품작 폴더에 따로 복사된 중복 MP4… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/LearningChat_ai_video_production.ai-video-cinematography-promptsfree-wavespeed-ai-video-api-nexaapi
Free WaveSpeed AI Video API — No Credit Card, Generate Videos Free via NexaAPI
See README.md for the full tutorial.
Links
🌐 NexaAPI: https://nexa-api.com
🔑 Free API Key: https://rapidapi.com/user/nexaquency
🐍 Python SDK: https://pypi.org/project/nexaapi/
📦 Node.js SDK: https://npmjs.com/package/nexaapi
