datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qualcomm-interactive-video-dataset
Dataset Card for Qualcomm Interactive Video Dataset
This is a FiftyOne dataset with 2900 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Voxel51/qualcomm-interactive-video-dataset")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/qualcomm-interactive-video-dataset.klingai-videos
KlingAI Video
This dataset contains 10394 video samples extracted from the nyuuzyou/klingai dataset.
Dataset Structure
Media files: video files in videos/ directory
Metadata: video_metadata.parquet contains all metadata including:
Original resource URLs
Dimensions (width, height)
Duration (for videos)
All other fields from the source dataset
Source
Original dataset: nyuuzyou/klingai
Processed and uploaded by the bitmind team for use in AIGC detection… See the full description on the dataset page: https://huggingface.co/datasets/bitmind/klingai-videos.selfie_and_video
Selfies and video dataset
4000 people in this dataset. Each person took a selfie on a webcam, took a selfie on a mobile phone. In addition, people recorded video from the phone and from the webcam, on which they pronounced a given set of numbers.
Includes folders corresponding to people in the dataset. Each folder includes 8 files (4 images and 4 videos).
💴 For Commercial Usage: To discuss your requirements, learn about the price and buy the dataset, leave a request on… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/selfie_and_video.2D_Video_Game_Cartoon_Character_Sprite-Sheets
Dataset Card for Dataset Name
Dataset Details
Experimental composition of 76 cartoon art-style video game character spritesheets. Resized to 512x512, mixed variation of animation styles.
Dataset Description
All images editted using Tiled image editting software as most assets are typically downloaded individually and not in sequence. I compiled each animation sequence into one img to display animations frame-by-frame evenly distributed across some common… See the full description on the dataset page: https://huggingface.co/datasets/mgane/2D_Video_Game_Cartoon_Character_Sprite-Sheets.video-dataset-audio_dataset
Video Dataset - audio_dataset
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips. Combined from tasks: task06, task07, task08
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-audio_dataset.Optimized_Video_Facial_Landmarks
Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset
Dataset Description
This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files.
License: The extracted feature data in this Parquet file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/PSewmuthu/Optimized_Video_Facial_Landmarks.Video-Moderation-4225
Video Moderation 4225
This is the fully materialized cleaned dataset used by
March-77/video-moderation-vlm
to train a Qwen3-VL-2B binary content-moderation adapter.
Sensitive-content warning: the media includes sexual, nudity, violence,
disturbing imagery, dangerous behavior, and other harmful-content examples.
Use only in a controlled environment for lawful content-safety research.
The project maintainer states that permission was obtained from the original
authors to… See the full description on the dataset page: https://huggingface.co/datasets/helloworldzzr/Video-Moderation-4225.hispanic-people-liveness-detection-video-dataset
Biometric Attack Dataset, Hispanic People
The similar dataset that includes all ethnicities - Anti Spoofing Real Dataset
The dataset for face anti spoofing and face recognition includes images and videos of hispanic people. 32,600+ photos & video of 16,300 people from 20 countries. The dataset helps in enchancing the performance of the model by providing wider range of data for a specific ethnic group.
The videos were gathered by capturing faces of genuine individuals… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/hispanic-people-liveness-detection-video-dataset.Emotion_Video_Facial_Landmarks
Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset
Dataset Description
This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files.
License: The extracted feature data in this CSV file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/PSewmuthu/Emotion_Video_Facial_Landmarks.asian-people-liveness-detection-video-dataset
Biometric Attack Dataset, Asian People
The similar dataset that includes all ethnicities - Anti Spoofing Real Dataset
The dataset for face anti spoofing and face recognition includes images and videos of asian people. 30,600+ photos & video of 15,300 people from 32 countries. All people presented in the dataset are South Asian, East Asian or Middle Asian. The dataset helps in enchancing the performance of the model by providing wider range of data for a specific ethnic… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/asian-people-liveness-detection-video-dataset.video_tagsThe MNIST dataset consists of 70,000 28x28 black-and-white images in 10 classes (one for each digits), with 7,000
images per class. There are 60,000 training images and 10,000 test images.black-people-liveness-detection-video-dataset
Biometric Attack Dataset, Black People
The similar dataset that includes all ethnicities - Anti Spoofing Real Dataset
The dataset for face anti spoofing and face recognition includes images and videos of black people. The dataset helps in enchancing the performance of the model by providing wider range of data for a specific ethnic group.
The videos were gathered by capturing faces of genuine individuals presenting spoofs, using facial presentations. Our dataset proposes… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/black-people-liveness-detection-video-dataset.cars-video-object-trackingThe collection of overhead video frames, capturing various types of vehicles
traversing a roadway. The dataset inculdes light vehicles (cars) and
heavy vehicles (minivan).video-dataset-new-combined_dataset
Video Dataset - combined_dataset
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips. Combined from tasks: task01, task02, task03, task04, task05
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-new-combined_dataset.video-imagesselfie-and-video-on-back-cameraThe dataset consists of selfies and video of real people made on a back camera
of the smartphone. The dataset solves tasks in the field of anti-spoofing and
it is useful for buisness and safety systems.Emotion_Video_Facial_Landmarks
Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset
Dataset Description
This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files.
License: The extracted feature data in this CSV file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/mac26/Emotion_Video_Facial_Landmarks.Anti-Spoofing-Real-Videos
Face Anti Spoofing Dataset - 98 000+ files
Dataset features 98,000+ files of real photos and videos of people from 170+ countries, representing 70,000+ unique individuals. By leveraging this dataset, developers can enhance spoofing detection techniques, improve recognition systems, and deploy anti-spoofing algorithms capable of preventing fraud in deep learning-based solutions.- Get the data
Dataset characteristics:
Characteristic
Data
Description
Live… See the full description on the dataset page: https://huggingface.co/datasets/ud-biometrics/Anti-Spoofing-Real-Videos.gelsight-mini-pretrain-video
GelSight Mini Pretrain · Video / Sequence Subset
🎬 Companion to yxma/gelsight-mini-pretrain.
Where the main repo treats every kept frame as an independent image, this repo
preserves temporal sequences — one row per frame, ordered, with explicit
sequence-id + position metadata, for video tactile pretraining.
Why this repo
The main repo's pipeline applies perceptual-hash dedupe within each capture
to drop near-identical adjacent frames. That's great for image-level… See the full description on the dataset page: https://huggingface.co/datasets/yxma/gelsight-mini-pretrain-video.video-dataset-task-02
Video Dataset - task-02
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv — mapping… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-task-02.PairedMNISTSimple toy dataset derived from MNIST, by shuffling the dataset to get image pairs. The task is to learn to get the multiplication value of the two figures in the two images.
The dataset is created mostly to learn how to upload datasets to huggingface
testing_qwen3vl_on_video
Dataset Card for harpreetsahota/random_short_videos
This is a FiftyOne dataset with 412 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/testing_qwen3vl_on_video")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/testing_qwen3vl_on_video.video-dataset-anonim_test
Video Dataset - anonim_test
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv — mapping… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-anonim_test.video-dataset-task02
Video Dataset - task02
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv — mapping… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-task02.video-dataset-combined_dataset
Video Dataset - combined_dataset
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv —… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-combined_dataset.video-dataset-test2040
Video Dataset - test2040
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames grouped by role (start, middle, end)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv —… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-test2040.fall_videos_GMNCSA24
Dataset Card for GMNCSA24-fall-frames-60pct-opencv2
This is a FiftyOne dataset with 76 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("pjramg/fall_videos_GMNCSA24")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/pjramg/fall_videos_GMNCSA24.video-dataset-pre1_test
Video Dataset - pre1_test
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames grouped by role (start, middle, end)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv —… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-pre1_test.video-dataset-new-combined_dataset-anonymized
Video Dataset - combined_dataset
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips. Combined from tasks: task01, task02, task03, task04, task05
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-new-combined_dataset-anonymized.video-dataset-C4KT-5XUY
Video Dataset - C4KT-5XUY
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames grouped by role (start, middle, end)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv —… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/video-dataset-C4KT-5XUY.
