rushilara/h2-video-detections
Video Detections and Query Clips Parquet outputs for the video detection and image semantic search pipeline (car-parts detector on video + RAV4 query images). Files video_detections.parquet — One row per video frame; each row has a list of object detections for that frame. query_longest_clips.parquet — One row per query image; each row has the longest contiguous video clip where the detected car parts appear, with a YouTube embed URL. Schema… See the full description on the dataset page: https://huggingface.co/datasets/rushilara/h2-video-detections.
Video Detections and Query Clips
Parquet outputs for the video detection and image semantic search pipeline (car-parts detector on video + RAV4 query images).
Files
- video_detections.parquet — One row per video frame; each row has a list of object detections for that frame.
- query_longest_clips.parquet — One row per query image; each row has the longest contiguous video clip where the detected car parts appear, with a YouTube embed URL.
Schema
video_detections.parquet
Each element in detections:
