video-summarization
Video_Summarization_For_Retail
Dataset Card for Video Summarization For Retail Dataset
This dataset contains short videos of shoppers in a retail setting along with the corresponding textual descriptions of each video.
Dataset Details
Curated by: Parker Lischwe
Language(s) (NLP): English
License: cc-by-sa-4.0
Uses
Navigate to Downloads directory where the zip file and python script have been downloaded to and run following commands in terminal:
pip install torch torchvision… See the full description on the dataset page: https://huggingface.co/datasets/Intel/Video_Summarization_For_Retail.nlp-summarization-audio-video
NLP Summarization Audio Video Data Notes
Dataset summary
This repository contains a preparation pipeline and a small metadata sample for NLP Summarization work with Audio Video inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated.
Included material
preprocess.py — loading, cleaning, and split preparation code.
dataset_infos.json — schema and split metadata.… See the full description on the dataset page: https://huggingface.co/datasets/Fabiookq0813/nlp-summarization-audio-video.nlp-summarization-audio-video
NLP Summarization Audio Video Data Notes
Dataset summary
This data card accompanies a lightweight NLP Summarization loader for Audio Video metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation.
Included material
prepare.py — loading, cleaning, and split preparation code.
dataset_infos.json — schema and split metadata.
metadata_sample.jsonl — small, human-readable records for checking the schema.… See the full description on the dataset page: https://huggingface.co/datasets/tiwalker/nlp-summarization-audio-video.nlp-summarization-video-text-v2
NLP Summarization Video Text Data Notes
Dataset summary
This data card accompanies a lightweight NLP Summarization loader for Video Text metadata. It is meant for pipeline inspection, source adaptation, and reproducible split preparation.
Included material
prepare.py — loading, cleaning, and split preparation code.
dataset_infos.json — schema and split metadata.
metadata_sample.jsonl — small, human-readable records for checking the schema.… See the full description on the dataset page: https://huggingface.co/datasets/daniel-ramos/nlp-summarization-video-text-v2.dataset_131467849_nlp_summarization_video_text
dataset_131467849_nlp_summarization_video_text.py
Dataset Summary
A nlp summarization dataset with video text modality, stored in huggingface format.
Preprocessing & Augmentation
Preprocessing: minimal
Augmentation: randaugment
Splits & Sampling
Split strategy: leave one out
Sampling: stratified
Quality & Labeling
Quality filtering: moderate
Labeling: pseudo label
Files… See the full description on the dataset page: https://huggingface.co/datasets/emakarovwell/dataset_131467849_nlp_summarization_video_text.dataset_131500778_nlp_summarization_video_text
dataset_131500778_nlp_summarization_video_text.py
Dataset Summary
A nlp summarization dataset with video text modality, stored in hdf5 format.
Preprocessing & Augmentation
Preprocessing: domain specific
Augmentation: mixup cutmix
Splits & Sampling
Split strategy: random 90 10
Sampling: contrastive
Quality & Labeling
Quality filtering: lenient
Labeling: self training
Files… See the full description on the dataset page: https://huggingface.co/datasets/rodrigomarques89/dataset_131500778_nlp_summarization_video_text.
