datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SAINetset_v8.0
SAINetset - Wildfire Smoke Detection Dataset
Dataset of real-world images captured by SAI (Sistema de Alerta de Incendios / Fire Alert System) surveillance nodes for wildfire smoke detection in Cordoba, Argentina.
Current version: v8.0 (January 2026)
About SAI
The SAI (Fire Alert System) is an open-source early wildfire detection platform developed by AlterMundi, a civil association in Argentina. The system uses distributed camera nodes with YOLO-based AI (powered by… See the full description on the dataset page: https://huggingface.co/datasets/SAINetset/SAINetset_v8.0.vision-opd-vqa14k-fullimage-curriculum-v8
Vision-OPD VQA14K Full-Image Curriculum v8
Private single-image visual-question-answering dataset.
Split
Rows
Train
14,000
Diagnostic validation
609
The repository contains 14,609 content-addressed media files (4,580,273,467 bytes). Paths in both Parquet files are relative to the repository root and follow media/<sha256-prefix>/<filename>.
from pathlib import Path
import pyarrow.parquet as pq
from huggingface_hub import snapshot_download
root =… See the full description on the dataset page: https://huggingface.co/datasets/yyy051007/vision-opd-vqa14k-fullimage-curriculum-v8.TR-HASH-Vision-v8-Demo-Images
TR-HASH Vision v8 demo images
This repository contains a fixed bank of 1,000 COCO 2017 validation images used by the public TR-HASH Vision v8 inference demo.
Source dataset: detection-datasets/coco, validation split
Images: 1,000
Purpose: reproducible public demo inputs
Metadata: manifest.json
The original image terms and attribution remain governed by the COCO source metadata and terms of use.
ai2thor-perspective-qa-800-qa-v8-splitsngld-grape-leaf-vlm-w-img-without-diff-ref-v8yolino-ttpla-benchmark
YOLinO TTPLA Benchmark (512×512)
TTPLA aerial imagery prepared for CAPSTONE / YOLinO polyline training and evaluation.
Each sample is a 512×512 PNG tile with matching NumPy (.npy) polyline labels.
Dataset structure
.
├── images/
│ ├── train/ # 905 PNG tiles
│ ├── val/ # 109 PNG tiles
│ └── test/ # 220 PNG tiles
└── labels/
├── train/ # 905 .npy label files
├── val/ # 109 .npy label files
└── test/ # 220 .npy label files
Image… See the full description on the dataset page: https://huggingface.co/datasets/V8heart/yolino-ttpla-benchmark.twitter-Shahaja40072599-2025.10.10-1976583196356870212-GVFSsQRMsf-V8RqV-part1RBY1_human_data_0401_v8_egoengineThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "rby1",
"total_episodes": 93,
"total_frames": 13116,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": [
82,
51,
52,
88,
78,
65,
33,
8,
37… See the full description on the dataset page: https://huggingface.co/datasets/Daniel233/RBY1_human_data_0401_v8_egoengine.synth-bg-remove-v8ttpla-yolino-1024
TTPLA YOLinO 1024×1024 Tiles
Dual-crop 1024² TTPLA tiles for backbone/predictor ablations (exp05–exp19) and full-resolution ISQ evaluation.
Layout
images/{train,val,test}/*.png
labels/{train,val,test}/*.npy
Each .npy label uses format version 3: polylines + instance_ids per wire.
Splits
Split
Images
train
1,810
val
218
test
440
Related
Code: V8heart/CAPSTONE
512² main benchmark: V8heart/yolino-ttpla-benchmark… See the full description on the dataset page: https://huggingface.co/datasets/V8heart/ttpla-yolino-1024.yolo_v8Target-PLP
Target-PLP
Target-PLP is a cross-domain test-only dataset for zero-shot evaluation of aerial power-line detection models.
It contains 194 aerial frames annotated with polyline instance labels in YOLinO format. Models are trained on TTPLA and evaluated on Target-PLP without fine-tuning.
Compared to TTPLA, Target-PLP is more complex and challenging: denser wire layouts, more parallel instances per frame, greater scene variation, and a different imaging domain. Annotations are… See the full description on the dataset page: https://huggingface.co/datasets/V8heart/Target-PLP.satellite_Roofs_MASK_train_valid_v8CoCount-train-aug-v8zxczczcx-v8
zxczczcx
A collection of question-and-answer pairs offering practical advice on personal, financial, and work-related topics. The dataset includes diverse scenarios and thoughtful, empathetic responses.
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
Quality of Remastered Dataset
The final quality is B, with a relative quality improvement of 46.7%.
Domain
Career-workplace (33%)
Personal-finance (33%)
Hr (17%)… See the full description on the dataset page: https://huggingface.co/datasets/NikitaSirotkin/zxczczcx-v8.shotpath-boundary-cot-v8-existing-full-core-ablation-data-20260728
ShotPath Boundary CoT v8 Existing Full-Core Ablation
Status: training candidate
This experiment tests whether explicit visible reasoning improves the canonical
open photography diagnosis task when no new teacher labels are added.
Task
Both variants use the canonical full-review prompt from
D:\data\shotpath\SHOTPATH_TASK_AND_DATA_CONTRACT.md. Directed rows add only
the allowed dimension scope line.
Data
1,532 directed exposure rows drawn from… See the full description on the dataset page: https://huggingface.co/datasets/purefall/shotpath-boundary-cot-v8-existing-full-core-ablation-data-20260728.life-advice-qa-v8
life_advice_qa
A dataset of question-answer pairs offering practical advice on personal, financial, and professional life challenges. Questions cover topics such as career changes, financial decisions, workplace etiquette, and wealth management. Answers are informative, empathetic, and focused on actionable solutions.
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
Quality of Remastered Dataset
The final quality is A, with a… See the full description on the dataset page: https://huggingface.co/datasets/NikitaSirotkin/life-advice-qa-v8.nayana-beir-eval-multilang_v8golf_mug_dataset_v8_max_min_cat_label_idw9_cord_v8juggernaut-XL-v8_prompt6urdu_images_doc_tags_all_v8juggernaut-XL-v8_prompt5my_first_lora_v89-datasetvlm-project-with-images-with-bbox-images-v8twitter-yaowo22kkowoi-2026.02.27-2027391183002173446-V8SfSVICQ9RF9vEg-part1masked_background_v8my_first_lora_v8-dataset
