datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
human-action-recognition-vl-enriched
Visualize on Visual Layer
Human-Action-Recognition-VL-Enriched
An enriched version of the Human-Action-Recognition Dataset with image captions, bounding boxes, and label issues! With this additional information, the Human-Action-Recognition dataset can be extended to various tasks such as image retrieval or visual question answering.
The label issues help curate a cleaner and leaner dataset.
Description
The dataset consists of 4 columns:
image_uri: The… See the full description on the dataset page: https://huggingface.co/datasets/visual-layer/human-action-recognition-vl-enriched.Audio-Visual-Speech-Recognition-VI
