CoolFace
20 results

video-dataset

nvidia /video-to-data-robot-dexterity-task-library-and-dataset Video to Data: Robot Dexterity Task Library and Dataset Dataset Description This dataset contains samples of human demonstrations on manipulation tasks retargeted to bimanual Sharpa robot hands and episodes of robot executions that mimic the original human demonstrations. The former allows a Video to Data user to easily experiment with the Video to Data grounding pipeline, and the latter is an example of the grounded robot data that can be generated with the Video… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/video-to-data-robot-dexterity-task-library-and-dataset.robotics12 likes18k downloads1mo agoHugging Facenetflix /Vera-Layered-Video-Dataset Dataset for Vera: A Layered Diffusion Model for Content-Preserving Video Editing Hongkai Zheng¹²* &nbsp;·&nbsp; Ta-Ying Cheng² &nbsp;·&nbsp; Benjamin Klein² &nbsp;·&nbsp; Yisong Yue¹ &nbsp;·&nbsp; Zhuoning Yuan²† ¹California Institute of Technology &nbsp;&nbsp; ²Netflix, Inc. *Work done during an internship at Netflix &nbsp; †Project Lead TL;DR: A layered diffusion framework for video editing. Vera jointly generates an edit layer, an alpha… See the full description on the dataset page: https://huggingface.co/datasets/netflix/Vera-Layered-Video-Dataset.videotext-to-video10K<n<100K59 likes13k downloads2mo agoHugging FaceElectronicHug /short_video_ocr_dataset Short Video OCR / ASR Dataset An actively curated research dataset for building OCR, ASR, subtitle-alignment, and video-transcript pipelines for short social videos. It combines source videos and extracted frames with human review artifacts and model-generated text candidates. The primary languages are Ukrainian and Russian; English or mixed-language content may also occur. Status: work in progress. Model outputs and pseudo-label candidates are not ground truth. Only… See the full description on the dataset page: https://huggingface.co/datasets/ElectronicHug/short_video_ocr_dataset.imageimage-to-text1K<n<10K0 likes11k downloads16h agoHugging Faceminkyuchoi /Temporal-Logic-Video-Dataset Temporal Logic Video (TLV) Dataset Temporal Logic Video (TLV) Dataset Synthetic and real video dataset with temporal logic annotation Explore the GitHub » NSVS-TL Project Webpage · NSVS-TL Source Code Overview The Temporal Logic Video (TLV) Dataset addresses the scarcity of state-of-the-art video datasets for long-horizon, temporally extended activity and object detection. It comprises two main components: Synthetic… See the full description on the dataset page: https://huggingface.co/datasets/minkyuchoi/Temporal-Logic-Video-Dataset.tabularquestion-answeringn<1K1 likes2.4k downloads2y agoHugging FaceVoxel51 /qualcomm-exercise-video-dataset-benchmark Dataset Card for Qualcomm Exercise Video Dataset (Benchmark) This is the benchmark split of the dataset as described here This is a FiftyOne dataset with 74 samples. Installation If you haven't already, install FiftyOne: pip install -U fiftyone Usage import fiftyone as fo from fiftyone.utils.huggingface import load_from_hub # Load the dataset # Note: other available arguments include 'max_samples', etc dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/qualcomm-exercise-video-dataset-benchmark.videon<1K3 likes2k downloads10mo agoHugging FaceTWTom /Letter_Vibration_Interference_Video_Dataset Dataset Card for Letter Vibration Interference Video Data Dataset Summary This dataset is collected using a 1920x1080 camera running at 60fps. It records the interference pattern generated by a Michelson Interferometer. The Interferometer is very sensitive to vibration, so while different vibration modes are performed, unique inference patterns are shown. We introduce vibration to the system by hand writing letters on the table which the interferometer is set up on.… See the full description on the dataset page: https://huggingface.co/datasets/TWTom/Letter_Vibration_Interference_Video_Dataset.1 likes1.9k downloads3y agoHugging Face