datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
amz-image-annotationspokemon-cards-image-and-annotationsmultilingual-image-annotations
Multilingual Image Annotations
Image annotations across 7 languages (en, es, fr, hi, zh, ar, pt) generated by google/gemma-4-31B-it via the Hugging Face Router. Each row pairs an image with an English description, multilingual descriptions, 21 VQA pairs (3 per language), and conditional object detections with normalized bounding boxes. When detections are present, a derivative image with rectangles drawn is included as boxed_image.
Stats
Images: 464… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/multilingual-image-annotations.pick_and_place_annotation_imageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "Unitree_G1_Inspire",
"total_episodes": 1035,
"total_frames": 322070,
"total_tasks": 14,
"total_videos": 2070,
"total_chunks": 2,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1035"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/KGB0/pick_and_place_annotation_image.
