datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
potter-plant-identificationRepaired_videoshhi-assist
HHI-Assist A Dataset and Benchmark of Human-Human Interaction in Physical Assistance Scenario
Saeed Saadatnejad, Reyhaneh Hosseininejad, Jose Barreiros, Katherine M. Tsui and Alexandre Alahi
https://ieeexplore.ieee.org/document/11071897
[webpage]
trilliongame
Bangumi Image Base of Trillion Game
This is the image base of bangumi Trillion Game, we detected 100 characters, 11831 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is the… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/trilliongame.TripVVT-10K
TripVVT-10K Dataset
News
2026.06: TripVVT has been accepted by ECCV 2026.
2026.04: The TripVVT paper is available on arXiv.
The project page is available at https://shaodingbao.github.io/TripVVT/.
TripVVT-10K is a large-scale dataset for in-the-wild Video Virtual Try-On (VVT). It contains 10,031 high-quality video samples with triplet supervision, covering upper-body garments, lower-body garments, and dresses.
TripVVT-10K is released together with the… See the full description on the dataset page: https://huggingface.co/datasets/TripVVT/TripVVT-10K.FraudBenchphisat2-s2-lightglue-triplets
PhiSat-2 / Sentinel-2 LightGlue Triplets
This dataset contains finalized strict triplet outputs generated from the local
PhiSat-2/Sentinel-2 LightGlue pipeline. Each accepted patch includes real
PhiSat-2, Sentinel-2 L1C, simulated PhiSat-2, OmniCloudMask, ESA WorldCover, and
Koppen-Geiger metadata.
Quality policy: fail closed. Patches are accepted only when registration,
geometry, nodata, PhiSat-2 cloud, WorldCover, and Koppen gates pass.
Current upload:
finalized pairs: 1… See the full description on the dataset page: https://huggingface.co/datasets/ESA-philab/phisat2-s2-lightglue-triplets.hilti-trimble-slam-challenge-2026
Hilti x Trimble SLAM Challenge 2026
The Hilti x Trimble SLAM Challenge 2026 dataset is a real-world robotics benchmark for evaluating visual-inertial SLAM and localization systems on active construction sites.
The dataset combines synchronized dual-fisheye imagery and inertial measurements with building floor plan priors and LiDAR-derived reference trajectories. It was created through a collaboration between Hilti, Trimble, and the Dynamic Robot Systems Group at the University… See the full description on the dataset page: https://huggingface.co/datasets/Hilti-Research/hilti-trimble-slam-challenge-2026.platonic-embeddingsencoder-decoder-trial-stat
Encoder/decoder trial: encoder-marginal report
Dataset: G-reen/encoder-decoder-trial-stat
Rows analysed: 122,933 (every kept (encoder, decoder, source row) triple; source G-reen/cc-re-2021-filtered shard 0, 2000 rows of at most 4000 words)
Prompt file: prompts/indirect_reference_dataset_train.json (turn 0 encodes the document, turn 1 reconstructs it from the encoding alone)
Encoders: 9 (granite-4.2-30b-nvfp4 [0], Ornith-1.5-35B-A3B-NVFP4 [1], Llama-3.3-70B-Instruct-NVFP4 [2]… See the full description on the dataset page: https://huggingface.co/datasets/G-reen/encoder-decoder-trial-stat.PPS_assets
PPS IsaacLab Assets
Scene and object assets for the IsaacLab manipulation tasks bundled with
TritiumR/pps — a research codebase for
steering pi0 / pi0.5 VLA policies.
The pps repository ships code only; these assets are distributed here
because they are too large to commit. Unpack them into IsaacLab/assets/ and
the task configs (which build asset paths relative to __file__) will resolve.
Contents
Path
Size
Source
ArtVIP/Interactive_scene/
~2.1 GB… See the full description on the dataset page: https://huggingface.co/datasets/Tritiumac/PPS_assets.trial_dataset
Trial Dataset (VQA)
This dataset contains various configurations for Visual Question Answering (VQA) tasks involving tables, figures, and multiple-choice options.
Dataset Structure
The dataset is split into multiple configurations based on the complexity of the input (number of tables/figures) and the response type (MCQ or Constructed Response).
How to Load
You can load any specific configuration using the datasets library:
from datasets import load_dataset
#… See the full description on the dataset page: https://huggingface.co/datasets/Krishna5T/trial_dataset.trial-v0-20250313
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/team-wonders/trial-v0-20250313.TridisThis is the first version of the dataset derived from the corpora used for TRIDIS (Tria Digita Scribunt).
TRIDIS encompasses a series of Handwriting Text Recognition (HTR) models trained using semi-diplomatic transcriptions of medieval and early modern manuscripts.
The semi-diplomatic transcription approach involves resolving abbreviations found in the original manuscripts and normalizing Punctuation and Allographs.
The dataset contains approximately 4,000 pages of manuscripts and is… See the full description on the dataset page: https://huggingface.co/datasets/magistermilitum/Tridis.triples-objekts
tripleS Objekts Archive (HD Images, Motion MP4s & Full Metadata)
Complete digital collectible photocards (Objekts) dataset for the K-Pop girl group tripleS (Modhaus / COSMO app).
Dataset Summary
Total Objekts: 11,025 unique digital collectibles
Total Images: 21,340 high-resolution card scans (Front & Back)
Total Motion & Voice Videos: 125 animated card loops & voice message MP4 clips
Members: All 24 members (S1–S24) + Sub-units (AAA, KRE, Assemble24, etc.)… See the full description on the dataset page: https://huggingface.co/datasets/fadhilafif98/triples-objekts.temporaryTriALS
TriALS: Triphasic-Aided Liver Lesion Segmentation Benchmark in Non-Contrast CT
Same patient across all four CT phases (non-contrast, arterial, portal venous, delayed). Top row: raw images. Bottom row: lesion annotations. Many lesions are occult on non-contrast CT and only become conspicuous after contrast administration — this is the core diagnostic challenge TriALS targets.
TriALS is the first multi-centre benchmark for liver lesion segmentation in non-contrast CT (NCCT)… See the full description on the dataset page: https://huggingface.co/datasets/marwankefah/TriALS.rebus-dataset
|🔄 🚍| Re-Bus: A Large and Diverse Multimodal Benchmark for evaluating the ability of Vision-Language Models to understand Rebus Puzzles
Understanding Rebus Puzzles requires a variety of skills such as image recognition, cognitive skills, commonsense reasoning, and multi-step reasoning, making this a challenging task for current Vision-Language Models. In this paper, we present Re-Bus, a large and diverse benchmark of 1,333 English Rebus Puzzles containing different artistic… See the full description on the dataset page: https://huggingface.co/datasets/TrishanuDas/rebus-dataset.triveni-raw
📦 Pretraining Corpus
📊 Dataset Overview
This dataset combines data from two major sources—Vaani and Flickr30k—to support multilingual and multimodal model pretraining.
Source
Languages
Samples per Language
Total Samples
Vaani
Hindi, English, Hinglish
30,195
90,585
Flickr30k
Hindi, English, Hinglish
31,014
93,042
Total
—
—
183,627
📁 Dataset Sources
🗣️ Vaani Dataset
License: CC-BY-4.0
Description:
VAANI is an… See the full description on the dataset page: https://huggingface.co/datasets/LingoIITGN/triveni-raw.vimeo90k_triplet
Dataset Card for "vimeo90k_triplet"
More Information needed
TriggerBenchTRIG
Dataset for the paper: Trade-offs in Image Generation: How Do Different Dimensions Interact?
Paper: https://huggingface.co/papers/2507.22100
TRIG is a benchmark for studying trade-offs across multiple image generation dimensions. It contains three tasks:
text_to_image
image_editing
subject_driven
All three splits share the same schema:
data_id: sample id, such as IQ-R_IQ-A_1
prompt: prompt used for generation or editing
dimensions: evaluated dimension pair
dimension_prompt:… See the full description on the dataset page: https://huggingface.co/datasets/RISys-Lab/TRIG.trigunstampede
Bangumi Image Base of Trigun Stampede
This is the image base of bangumi Trigun Stampede, we detected 37 characters, 3814 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is the… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/trigunstampede.sample_datasetLogo-Recognition-ResNet50-TripletNet-Embeddings-DatasetCOCO2017-stuffTriveni
📦 Pretraining Corpus
📊 Dataset Overview
This dataset combines data from two major sources—Vaani and Flickr30k—to support multilingual and multimodal model pretraining.
Source
Languages
Samples per Language
Total Samples
Vaani
Hindi, English, Hinglish
30,195
90,585
Flickr30k
Hindi, English, Hinglish
31,014
93,042
Total
—
—
183,627
📁 Dataset Sources
🗣️ Vaani Dataset
License: CC-BY-4.0
Description:
VAANI is an… See the full description on the dataset page: https://huggingface.co/datasets/LingoIITGN/Triveni.puck_teleop_filtered_trimmed_fk_no_stateThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "widowx_ai",
"total_episodes": 183,
"total_frames": 17059,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:183"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/kobikelemen/puck_teleop_filtered_trimmed_fk_no_state.image_tripletspuck_teleop_filtered_trimmed_fkThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "widowx_ai",
"total_episodes": 183,
"total_frames": 17059,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:183"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/kobikelemen/puck_teleop_filtered_trimmed_fk.
