video-dataset
Create_dataset_from_videovideomae-base-finetuned-kinetics-finetuned-shoplifting-dataset-2videomae-base-finetuned-kinetics-finetuned-shoplifting-datasetvideomae-base-finetuned-fight-datasetvideomae-base-finetuned-kinetics-finetuned-traffic-dataset-maevideomae-base-finetuned-kinetics-finetuned-my-dataset-4-epochsVideoMAEF-finetuned-ARSL-diverse-datasetvideomae-base-finetuned-kinetics-finetuned-Accident-dataset
video-to-data-robot-dexterity-task-library-and-dataset
Video to Data: Robot Dexterity Task Library and Dataset
Dataset Description
This dataset contains samples of human demonstrations on manipulation tasks retargeted to bimanual Sharpa robot hands and episodes of robot executions that mimic the original human demonstrations. The former allows a Video to Data user to easily experiment with the Video to Data grounding pipeline, and the latter is an example of the grounded robot data that can be generated with the Video… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/video-to-data-robot-dexterity-task-library-and-dataset.Vera-Layered-Video-Dataset
Dataset for Vera: A Layered Diffusion Model for Content-Preserving Video Editing
Hongkai Zheng¹²* ·
Ta-Ying Cheng² ·
Benjamin Klein² ·
Yisong Yue¹ ·
Zhuoning Yuan²†
¹California Institute of Technology ²Netflix, Inc.
*Work done during an internship at Netflix †Project Lead
TL;DR: A layered diffusion framework for video editing. Vera jointly generates an edit layer, an alpha… See the full description on the dataset page: https://huggingface.co/datasets/netflix/Vera-Layered-Video-Dataset.short_video_ocr_dataset
Short Video OCR / ASR Dataset
An actively curated research dataset for building OCR, ASR, subtitle-alignment,
and video-transcript pipelines for short social videos. It combines source
videos and extracted frames with human review artifacts and model-generated
text candidates. The primary languages are Ukrainian and Russian; English or
mixed-language content may also occur.
Status: work in progress. Model outputs and pseudo-label candidates are
not ground truth. Only… See the full description on the dataset page: https://huggingface.co/datasets/ElectronicHug/short_video_ocr_dataset.Temporal-Logic-Video-Dataset
Temporal Logic Video (TLV) Dataset
Temporal Logic Video (TLV) Dataset
Synthetic and real video dataset with temporal logic annotation
Explore the GitHub »
NSVS-TL Project Webpage
·
NSVS-TL Source Code
Overview
The Temporal Logic Video (TLV) Dataset addresses the scarcity of state-of-the-art video datasets for long-horizon, temporally extended activity and object detection. It comprises two main components:
Synthetic… See the full description on the dataset page: https://huggingface.co/datasets/minkyuchoi/Temporal-Logic-Video-Dataset.qualcomm-exercise-video-dataset-benchmark
Dataset Card for Qualcomm Exercise Video Dataset (Benchmark)
This is the benchmark split of the dataset as described here
This is a FiftyOne dataset with 74 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/qualcomm-exercise-video-dataset-benchmark.Letter_Vibration_Interference_Video_Dataset
Dataset Card for Letter Vibration Interference Video Data
Dataset Summary
This dataset is collected using a 1920x1080 camera running at 60fps. It records the interference pattern generated by a Michelson Interferometer. The Interferometer is very sensitive to vibration, so while different vibration modes are performed, unique inference patterns are shown. We introduce vibration to the system by hand writing letters on the table which the interferometer is set up on.… See the full description on the dataset page: https://huggingface.co/datasets/TWTom/Letter_Vibration_Interference_Video_Dataset.
