CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01spongy /physical-ai-bench-step-30000-videos PhysicalAIBench Step 30000 Model Comparison This repository contains two filename-aligned sets of 5,220 MP4 outputs from step_30000 evaluation runs. The files are presented through the companion PhysicalAI Video Gallery. Model sets Directory Model Files Bytes videos/ DC-AE v0.2, Cosmos encoder + causal decoder 5,220 4,244,965,541 videos_wan22_vae/ Wan 2.2 VAE, phase 3 PDX 5,220 4,818,900,749 The two directories have an exact 1:1 basename match.… See the full description on the dataset page: https://huggingface.co/datasets/spongy/physical-ai-bench-step-30000-videos.video10K<n<100K0 likes1.9k downloads2mo agoHugging Face02chikuwa-AI /AudioSet_balanced_videovideo10K<n<100K0 likes750 downloads2mo agoHugging Face03meta-ai-for-media-research /movie_gen_video_bench Dataset Card for the Movie Gen Benchmark Movie Gen is a cast of foundation models that generates high-quality, 1080p HD videos with different aspect ratios and synchronized audio. Here, we introduce our evaluation benchmark "Movie Gen Bench Video Bench", as detailed in the Movie Gen technical report (Section 3.5.2). To enable fair and easy comparison to Movie Gen for future works on these evaluation benchmarks, we additionally release the non cherry-picked generated videos from… See the full description on the dataset page: https://huggingface.co/datasets/meta-ai-for-media-research/movie_gen_video_bench.text1K<n<10K29 likes685 downloads2y agoHugging Face04ai-api-key-free-finder /quantum-video Dataset Card for Dataset Name quantum suite video quantum suite video dataset, used to train the quantum suite video model. Dataset Details will upload, collecting data Dataset Description This dataset is aimed to be curated for the quantum suite video dataset with video from wikimedia (Cc allowing commercial use), this dataset allows commercial use, and will become useful for you to use in your video models, the size is aimed to be 2.5tb, after we… See the full description on the dataset page: https://huggingface.co/datasets/ai-api-key-free-finder/quantum-video.text-to-video0 likes370 downloads27d agoHugging Face05ymo2017 /ai-shorts-videosvideon<1K0 likes362 downloads1d agoHugging Face06nafisatibrahim /wat.ai-so101-videosimagen<1K0 likes311 downloads2mo agoHugging Face07stepfun-ai /Step-Video-T2V-EvalThis dataset contains the data of the paper Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model. Code: https://github.com/stepfun-ai/Step-Video-T2V Project page: https://yuewen.cn/videos videotext-to-videon<1K1 likes278 downloads2y agoHugging Face085CD-AI /Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated Dataset Card for 5CD-AI/Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated This translated dataset includes: LLaVA-Video-178K: 178,509 caption entries, 960,791 open-ended QA (question and answer) items, and 196,198 multiple-choice QA items. The video source of the original dataset is in this repo: lmms-lab/LLaVA-Video-178K textvisual-question-answering1M<n<10M1 likes273 downloads2y agoHugging Face09These-Guys-Know /tgk-ai-video-generators-2026 Permanent dataset archive: https://doi.org/10.5281/zenodo.22703594 Five AI Video Generators Tested on Dialogue, Action and an Advert These Guys Know tested Seedance 2.5, MiniMax H3, FLUX 3 Video, Gemini Omni 1.1 Flash and HappyHorse 1.1 on 1 September 2026. Every model received the same three ten-second, 16:9 text-to-video tasks: a father interrupting a computer game, a three-person fight inside a fixed hotel lobby and a Mango Cola advert with an exact product name. We retained… See the full description on the dataset page: https://huggingface.co/datasets/These-Guys-Know/tgk-ai-video-generators-2026.tabulartext-to-videon<1K0 likes201 downloads12d agoHugging Face10sosa123454321 /videobridge-ai-kb 🎬 VideoBridge AI — Knowledge Base (RAG) The single knowledge base for the VideoBridge AI project — used by the Telegram bot (@VideoBrige_bot) and the site panel for RAG search (/ai/rag). This is the only Hugging Face dataset created for this project (no duplicate models/spaces/databases). 🔄 Auto-update A daily scraper (scraper/kb_scrape.py) refreshes this dataset with: 📚 Academic papers (arXiv): text-to-video, video generation, video diffusion, RAG 📰 AI/tech… See the full description on the dataset page: https://huggingface.co/datasets/sosa123454321/videobridge-ai-kb.question-answering0 likes174 downloads11d agoHugging Face11Juasmo /AI-Vault-Videostextn<1K0 likes149 downloads4d agoHugging Face12HelioAI /Old-AI-Video 📁 Old AI Video — Архив оригинальных AI-видео 🔒 Правообладатель Абдулаев Самад Германович 📧 usnul.noxil@gmail.com 📋 Что это Данный датасет является полным эталонным архивом всех оригинальных AI-видео и связанных материалов, созданных Абдулаевым Самадом Германовичем в период с 21 декабря 2024 года и опубликованных на Telegram-канале «ИИ-видео» (Telegram ID: -1002330425412). Архив создан для: ✅ Фиксации авторства — доказательство того, что все материалы… See the full description on the dataset page: https://huggingface.co/datasets/HelioAI/Old-AI-Video.n<1K0 likes143 downloads6mo agoHugging Face13HiDream-ai /VIP-200K-Videogated VIP-200K-Video Overview This repository contains only the video files (.tar archives) from the VIP-200K dataset. For the complete dataset (including JSON annotations, face frames, and segment metadata), please download HiDream-ai/VIP-200K. File Structure video/ ├── vip200k_train_0001_of_0100.tar ├── vip200k_train_0002_of_0100.tar ├── ... └── vip200k_train_0100_of_0100.tar Each .tar archive contains video clips organized by YouTube video ID: {video_id}/… See the full description on the dataset page: https://huggingface.co/datasets/HiDream-ai/VIP-200K-Video.text-to-video3 likes131 downloads6mo agoHugging Face14yueying-117 /ai-generated_videotextn<1K0 likes125 downloads11mo agoHugging Face15Memories-ai /UGC-VideoCap UGC-VideoCaptioner Dataset Real-world user-generated videos, especially on platforms like TikTok, often feature rich and intertwined audio-visual content. However, existing video captioning benchmarks and models remain predominantly visual-centric, overlooking the crucial role of audio in conveying scene dynamics, speaker intent, and narrative context. This lack of full-modality datasets and lightweight, capable models hampers progress in fine-grained, multimodal video… See the full description on the dataset page: https://huggingface.co/datasets/Memories-ai/UGC-VideoCap.video-text-to-text1 likes89 downloads1y agoHugging Face16AlwaysDave /AI-Vault-Videostextn<1K0 likes77 downloads1mo agoHugging Face17rithwikn /ai_cctv_videosvideon<1K0 likes70 downloads8mo agoHugging Face18meta-ai-for-media-research /movie_gen_video_bench_no_generations Dataset Summary Please see the full dataset huggingface page text1K<n<10K9 likes66 downloads2y agoHugging Face19hellohihiloy789 /ai-translate-video Video Localization Timing Dataset Dataset Summary This dataset is a self-authored synthetic benchmark for multilingual video localization workflows. It focuses on the timing pressure that appears when subtitle segments are translated across languages and then reviewed for dubbing fit, subtitle-window preservation, and lip-sync risk. The package is designed for repository-safe experimentation and documentation. It does not contain third-party video, third-party audio… See the full description on the dataset page: https://huggingface.co/datasets/hellohihiloy789/ai-translate-video.documenttranslation1K<n<10K0 likes62 downloads4mo agoHugging Face20chikuwa-AI /AudioSet_unbalanced_videovideo100K<n<1M0 likes60 downloads2mo agoHugging Face21hg-ai /video_upscale_engine0 likes56 downloads2y agoHugging Face22Automation-Tribe /Telugu-NLP-AI-Dialect-Comedy-video-DatasetTelugu is one of the sweetest and oldest languages of India. A deep Dive into Telugu its spoken in 2 states and majorly 16 regional dailects. This Dataset help you perform operations in NLP and Speech Recognition Models towards telugu Dialects. text-classification1 likes52 downloads2y agoHugging Face23ankamma12 /Telugu-NLP-AI-Dialect-Comedy-video-DatasetTelugu is one of the sweetest and oldest languages of India. A deep Dive into Telugu its spoken in 2 states and majorly 16 regional dailects. This Dataset help you perform operations in NLP and Speech Recognition Models towards telugu Dialects. text-classification0 likes52 downloads6mo agoHugging Face24psdn-ai /egocentric-activity-video-samplesgated Egocentric Activity Video Samples This sample shows first-person activity video for reviewing task flow, camera perspective, and real-world action structure before scoping a larger delivery. What This Shows Egocentric footage of everyday task activity Clip-level metadata for task and scene review A view of capture quality, framing, and movement patterns Dataset Specifications Field Value Modality Video Domain First-person activity… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/egocentric-activity-video-samples.tabularvideo-classificationn<1K0 likes42 downloads1mo agoHugging Face25jbilcke-hf /ai-video-model-benchmarkvideon<1K0 likes37 downloads2y agoHugging Face26soradata /ai_videos_Financevideon<1K0 likes36 downloads2mo agoHugging Face2734data /kaggle-ai-videovideon<1K0 likes35 downloads2mo agoHugging Face28K-University-AIED /LearningChat_ai_video_production Hallym AI Video Production Practice 2025-2 Public Dataset 1. 데이터셋 개요 데이터셋명: 한림대학교 AI영상제작실습 2025-2 공개용 데이터셋 교과목명: AI영상제작실습 학기: 2025-2 생성 배경: 2025학년도 2학기 AI영상제작실습 수업에서 조별로 제작·제출한 AI 기반 영상 결과물을 공개용 데이터셋 형태로 정리한 것이다. 목적: 수업 기반 AI 영상 창작 결과물을 공개 아카이브 형태로 정리하고, 작품 단위 메타데이터를 함께 제공하기 위함이다. 2. 데이터셋 범위 총 작품 수: 20편 데이터 단위: 조별 제출 영상 1편 = metadata.csv 1행 포함 대상: 1조부터 20조까지 각 팀 폴더의 원본 MP4 1개 제외 대상: 보고서 파일(.pdf, .docx, .hwp) 라이선스 동의서 파일 AI제작콘텐츠 발표회 2025 출품작 폴더에 따로 복사된 중복 MP4… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/LearningChat_ai_video_production.documentn<1K0 likes34 downloads6mo agoHugging Face29gunahkarcasper /ai-video-cinematography-promptstextn<1K0 likes33 downloads8mo agoHugging Face30nickyni /free-wavespeed-ai-video-api-nexaapi Free WaveSpeed AI Video API — No Credit Card, Generate Videos Free via NexaAPI See README.md for the full tutorial. Links 🌐 NexaAPI: https://nexa-api.com 🔑 Free API Key: https://rapidapi.com/user/nexaquency 🐍 Python SDK: https://pypi.org/project/nexaapi/ 📦 Node.js SDK: https://npmjs.com/package/nexaapi 2 likes31 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.