datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
panlex-meanings
Dataset Card for panlex-meanings
This is a dataset of words in several thousand languages, extracted from https://panlex.org.
Dataset Details
Dataset Description
This dataset has been extracted from https://panlex.org (the 20240301 database dump) and rearranged on the per-language basis.
Each language subset consists of expressions (words and phrases).
Each expression is associated with some meanings (if there is more than one meaning, they are in separate… See the full description on the dataset page: https://huggingface.co/datasets/gtak1/panlex-meanings.gta-data-files-universalgta-data-files-universalgta-data-files-universalgta-files-auroraGTA5self-driving-GTA-V
Self Driving GTA V Dataset
Dataset Varients
Mini : Link
Training Data(1-100) : Link
Training Data(101-200) : Link
Info
Image Resolution : 270, 480
Mode : RGB
Dimension : (270, 480, 3)
File Count : 100
Size : 1.81 GB/file
Total Data Size : 362 GB
Total Frames : 1 Million
Data Set sizes
Mini :
Folder Name : mini
Files : 01
Total Size : 1.81 GB
Total Frames : 5000
First Half
Folder Name : Training Data(1-100)
Files :… See the full description on the dataset page: https://huggingface.co/datasets/sartajbhuvaji/self-driving-GTA-V.GTA-Human
Playing for 3D Human Recovery (TPAMI 2024)
Homepage
Toolbox
Paper
Updates
[2024-10-02] GTA-Human datasets are now available on HuggingFace!
[2024-09-19] Release of GTA-Human II Dataset
[2022-07-08] Release of GTA-Human Dataset on MMHuman3D
Datasets
Please click on the dataset name for download links and visualization instructions.
Features
GTA-Human
GTA-Human II
Num of Scenes
20,005
10,224
Num of Person Sequences
20,005
35,352
Color Images… See the full description on the dataset page: https://huggingface.co/datasets/caizhongang/GTA-Human.easyr1-103k-4MP-jedi-ui-vision-gta1-data
easyr1-103k-4MP-jedi-ui-vision-gta1-data
Merged dataset composed of the following sources:
datasets/easyr1-63k-nores-jedi-fix-synced-ui-vision-manually-labeled-icon-data-from-yt-4MP (63031 samples in split train)
datasets/easyr1-grounding-gta1-4MP-easy-qwen7b-hard-gta1-7b (39943 samples in split train)
Summary
Generated on: 2025-09-18 06:29:16 UTC
Split: train
Column strategy: intersection
Samples after merge: 102974
Usage
from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-103k-4MP-jedi-ui-vision-gta1-data.GTASAGTA
GTA: A Benchmark for General Tool Agents
[📃 Paper]
[🌐 Project Page]
[</> Code]
Dataset Summary
GTA is a benchmark to evaluate the tool-use capability of LLM-based agents in real-world scenarios. It features three main aspects:
Real user queries. The benchmark contains 229 human-written queries with simple real-world objectives but implicit tool-use, requiring the LLM to reason the suitable tools and plan the solution steps.
Real deployed tools.GTA… See the full description on the dataset page: https://huggingface.co/datasets/Jize1/GTA.gtasa-01
GTASA-01: Multi-Actor Video Corpus with Perfect Spatiotemporal Annotations
GTASA-01 is the sample corpus released with the ICLR 2026 Tiny Paper
GEST-Engine: Controllable Multi-Actor Video Synthesis with Perfect Spatiotemporal Annotations.
The corpus contains 398 procedurally generated multi-actor stories produced by the GEST-Engine,
each accompanied by a Graph of Events in Space and Time (GEST) specification, an engine-rendered
RGB video with dense spatiotemporal annotations, and —… See the full description on the dataset page: https://huggingface.co/datasets/nnc-001/gtasa-01.GTA-UAV-LR
GTA-UAV dataset
For more information, please check our project page.
Sources
Repository: https://github.com/Yux1angJi/GTA-UAV
Paper: https://arxiv.org/abs/2409.16925
gta-data-fulleasyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter-tokenizedgta5-cityscapes-labelingeasyr1-103k-4MP-jedi-ui-vision-gta1-data-bbox-filtered-0p05easyr1-grounding-gta1-4MP-easy-qwen7b-hard-gta1-7b
easyr1-grounding-gta1-4MP-easy-qwen7b-hard-gta1-7b
This dataset was generated from filtered GTA shards with images streamed from ZIP archives.
Generated on: 2025-09-18 05:15:30 UTC
Script: push_easyr1_zip_shards_to_hf.py
Filters directory: /p/project1/synthlaion/awadalla1/gta-grounding-data-filters
JSONL glob: gta_shard_*zip.jsonl
Resize max: 4.0 MP
Prompt format: gta1 (output: coordinates)
Random seed: 42
Deduplicate: False
Debug images: True
System Prompt
You are… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-grounding-gta1-4MP-easy-qwen7b-hard-gta1-7b.easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter-coord-grid
easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter-coord-grid
Augmented version of easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter with a fixed 100px coordinate grid overlay.
Each image is overlaid with vertical and horizontal grid lines every 100
pixels at native resolution. Major ticks (every 1 steps)
are emphasized and axis labels show pixel values to help models localize
precise coordinates.
Summary
Generated on:… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter-coord-grid.gta-vc-map
gta-vc-map
Mapas Vice City e Liberty City para o launcher SA-MP (fonte: SAxVCxLC, OpenVice, GTA:LC).
Layout (prefixo por variante):
vc/SAMP/gta.dat, vc/data/maps/VC/**, vc/texdb/vc_map.img, vc/texdb/vc_col.img
lc/SAMP/gta.dat, lc/data/maps/LC/**, lc/texdb/lc_map.img, lc/texdb/lc_col.img
mapas SA vem do dataset base (Soiza/gta-data-full) + OBB.
Tripoint_GTA_Data_Peasyr1-103k-4MP-jedi-ui-vision-gta1-data-sampling-not-all-correct-stage-one-temp-1_1-RL-a-keasyr1-103k-4MP-jedi-ui-vision-gta1-data-sampling-not-all-correct-stage-one-temp-1_1-RLCalliBench
🧠 CalliReader: Contextualizing Chinese Calligraphy via an Embedding-aligned Vision Language Model
📂 Code
📄 Paper
CalliBench is aimed to comprehensively evaluate VLMs' performance on the recognition and understanding of Chinese calligraphy.
📦 Dataset Summary
Samples: 3,192 image–annotation pairs
Tasks: Full-page recognition and Contextual VQA (choice of author/layout/style, bilingual interpretation, and intent analysis).
Annotations:
Metadata of author… See the full description on the dataset page: https://huggingface.co/datasets/gtang666/CalliBench.GTA5subset
GTA5 Subset for Zero-Shot Domain Adaptive Semantic Segmentation
This repository contains a curated subset of the GTA5 dataset, specifically designed for experiments in the paper Zero Shot Domain Adaptive Semantic Segmentation by Synthetic Data Generation and Progressive Adaptation. This subset includes images and labels necessary for training and evaluating models in zero-shot domain adaptive semantic segmentation scenarios.
The original GTA5 dataset was extracted from… See the full description on the dataset page: https://huggingface.co/datasets/roujin/GTA5subset.GTA-UAV-HR
GTA-UAV dataset
# Merge splited files
cat drone_part_* > drone.tar.gz
# Extract the archive
tar -xzvf drone.tar.gz
tar -xzvf satellite.tar.gz
For more information, please check our project page.
Sources
Repository: https://github.com/Yux1angJi/GTA-UAV
Paper: https://arxiv.org/abs/2409.16925
easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter
easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter
Augmented version of easyr1-57k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced with coordinate jitter.
For each original example, 1 additional copies were created. Each copy
randomly jitters the target coordinate by ±1 pixel in both X and Y. The
assistant coordinate in messages is updated, and bbox/normalized_bbox
are shifted when present.
Summary
Generated on: 2025-09-07 17:36:27 UTC
Source… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter.GTASAMBeasyr1-21k-jedi-grounding-4MP-gta1-nores-fixed
easyr1-21k-jedi-grounding-4MP-gta1-nores-fixed
This dataset is a fixed version of /Users/anasawadalla/Desktop/cua/easyr1-21k-jedi-grounding-4MP. It applies two changes:
Rebuilds prompts/messages to GTA1 format without resolution in the system prompt
Removes samples where a 100x100 patch around the bbox midpoint is a solid color
Summary
Generated on: 2025-09-05 23:26:46 UTC
Source dataset: /Users/anasawadalla/Desktop/cua/easyr1-21k-jedi-grounding-4MP
Split: train… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-21k-jedi-grounding-4MP-gta1-nores-fixed.GTA_data_no_web
