datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sam3d-flat-20260329-022951metaworld_mt10_gen_sam3_masksThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": null,
"total_episodes": 500,
"total_frames": 46327,
"total_tasks": 10,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 80,
"splits": {
"train": "0:500"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Beegbrain/metaworld_mt10_gen_sam3_masks.basil-segmentation-sam3
maximilian-franz/basil-segmentation-sam3
Per-instance segmented basil crops produced by the sam3 backend. This is a Hugging Face ImageFolder dataset: file_name points to the black-background masked crop used for downstream image analysis and bbox_file_name points to the corresponding unmasked rectangular crop. Empty masks are omitted. Bounding boxes use native source-frame coordinates.
Load it with:
from datasets import load_dataset
ds =… See the full description on the dataset page: https://huggingface.co/datasets/maximilian-franz/basil-segmentation-sam3.det_qwen3vl_sam3
注意
请解压DIOR.zip和FAIR1Mv2.zip,将解压得到的文件夹放到./DIOR路径下和./FAIR1Mv2路径下,再将DIOR_incontext.jsonl与DIOR_incontext.parquet放到./DIOR/conv路径下。
low-alt-satellite-image-dataset-5k-sam3-segmented_jsonlow-alt-satellite-image-dataset-5k-sam3-segmentedblackline-atlas-sam3-real-eval-v2
Blackline Atlas SAM3 Real-Image Eval Pack
This dataset packages the Blackline Atlas SAM3/SAM3.1 selected-site evidence eval pack
with real SimSat Sentinel image pairs.
It is an evaluation and integration dataset, not a training benchmark. The cases are exact
civilian lifeline sites with current/baseline satellite frames, text prompts, expected
visual evidence tags, expected triage action, and optional normalized bboxes.
Contents
Cases: 22
Images: 44 PNG files
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/ChrisRPL/blackline-atlas-sam3-real-eval-v2.sam3-low-dice-2d-nnunet
SAM3 low-Dice 2D datasets for nnU-Net
Private research export of two small 2D datasets on which the balanced-finish
SAM3 LoRA validation Dice was below 0.5. The purpose is to test whether a
dataset-specific nnU-Net can fit these data and to distinguish data/training
limitations from inference bugs.
Dataset
SAM3 Dice
SAM3 IoU
Evaluated validation images
Actual SAM3 training images
DRIVE
0.212233
0.118717
2
14
RAVIR
0.224709
0.128455
2
16
The two-image validation… See the full description on the dataset page: https://huggingface.co/datasets/MedicalSAM3/sam3-low-dice-2d-nnunet.person-benchmark-sam3
person-benchmark-sam3
Element: manak0/Detect-Person (person)
Classes: 0: person
Layout: yolo — 260 images, 13081 boxes
Source: tool:build
split
images
labeled
boxes
train
260
260
13081
Steps:
build 2026-09-16T16:32:11Z
person-new-cctv-sam3
Person New CCTV — SAM3 Annotated
YOLO-format person boxes for synthetic CCTV images from jjjlimaus/person-new-cctv-synthetic.
Annotator: SAM3 (prepare_dataset_with_sam3.py --mode person)
Class: person only (nc: 1)
Augmentation: original images plus horizontal flips (_aug_hflip), each flip re-annotated with SAM3
Samples: 2060 (train 1648 / val 412 / test 0)
Split: 8:2:0 (seed 42); flips stay in the same split as the source image
Resolution: all images stretched to 1024×1024… See the full description on the dataset page: https://huggingface.co/datasets/jjjlimaus/person-new-cctv-sam3.sam3-segment-test-wildlife
Image Segmentation: Deer using SAM3
This dataset contains semantic segmentation maps for deer segmented in images from davanstrien/ena24-detection using Meta's SAM3.
Generated using: uv-scripts/sam3 segmentation script
Statistics
Objects Segmented: deer
Total Instances: 0
Images with Detections: 0 / 5 (0.0%)
Average Instances per Image: 0.00
Output Format: semantic-mask
Processing Details
Source Dataset: davanstrien/ena24-detection
Model: facebook/sam3… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/sam3-segment-test-wildlife.traffi-sam3-inputperson-high-cctv-sam3
Person High CCTV SAM3
SAM3 --mode person YOLO annotations for jjjlimaus/person-high-cctv-synthetic.
Source: KIE nano-banana-2 bird's-eye / elevated person CCTV images (originals + hflip / l30 / r30)
Class: person (nc: 1)
Layout: Ultralytics {train,val,test}/{images,labels,visualizations}/ plus dataset.yaml
Labels: YOLO class cx cy w h (.txt) and SAM3 masks (.npz)
Split: random 80:12:8 from SAM3 stage-2 (_pick_split)
Counts: 824 paired samples (train 650, val 106, test 68).… See the full description on the dataset page: https://huggingface.co/datasets/jjjlimaus/person-high-cctv-sam3.traffi-sam3-resized-inputsam3-segment-test
Image Segmentation: Photograph using SAM3
This dataset contains semantic segmentation maps for photograph segmented in images from davanstrien/newspapers-with-images-after-photography using Meta's SAM3.
Generated using: uv-scripts/sam3 segmentation script
Statistics
Objects Segmented: photograph
Total Instances: 0
Images with Detections: 0 / 10 (0.0%)
Average Instances per Image: 0.00
Output Format: semantic-mask
Processing Details
Source… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/sam3-segment-test.sam3-cats-test
Object Detection: Cat Detection using sam3
This dataset contains object detection results (bounding boxes) for cat detected in images from cats_vs_dogs using Meta's SAM3 (Segment Anything Model 3).
Generated using: uv-scripts/sam3 detection script
Detection Statistics
Objects Detected: cat
Total Detections: 15
Images with Detections: 5 / 5 (100.0%)
Average Detections per Image: 3.00
Processing Details
Source Dataset: cats_vs_dogs
Model: facebook/sam3… See the full description on the dataset page: https://huggingface.co/datasets/felipehsilveira/sam3-cats-test.sam3-segment-test-debug
Image Segmentation: Animal using SAM3
This dataset contains semantic segmentation maps for animal segmented in images from davanstrien/ena24-detection using Meta's SAM3.
Generated using: uv-scripts/sam3 segmentation script
Statistics
Objects Segmented: animal
Total Instances: 8
Images with Detections: 5 / 5 (100.0%)
Average Instances per Image: 1.60
Output Format: semantic-mask
Processing Details
Source Dataset: davanstrien/ena24-detection
Model:… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/sam3-segment-test-debug.sam3-segment-deer-testperson-kie-cctv-sam3-1024sq
person-kie-cctv-sam3-1024sq
SAM3 (--mode person) YOLO labels for the 1024×1024 synthetic CCTV person images.
Source images: jjjlimaus/person-kie-cctv-synthetic-1024sq
Images kept: 2022
Label files: 2022
Layout: Ultralytics {train,val,test}/{images,labels} + dataset.yaml
Class: person (nc=1)
person-new-cctv-sam3-1024sq
person-new-cctv-sam3-1024sq
SAM3 (--mode person) YOLO labels for the 1024×1024 synthetic CCTV person images.
Source images: jjjlimaus/person-new-cctv-synthetic-1024sq
Images kept: 2062
Label files: 2062
Layout: Ultralytics {train,val,test}/{images,labels} + dataset.yaml
Class: person (nc=1)
pascal-parts-sam3low-alt-satellite-image-dataset-5k-sam3-segmented_json_vehiclessam3-test
Object Detection: Person Detection using sam3
This dataset contains object detection results (bounding boxes) for person detected in images from huggan/few-shot-art-painting using Meta's SAM3 (Segment Anything Model 3).
Generated using: uv-scripts/sam3 detection script
Detection Statistics
Objects Detected: person
Total Detections: 9
Images with Detections: 3 / 3 (100.0%)
Average Detections per Image: 3.00
Processing Details
Source Dataset:… See the full description on the dataset page: https://huggingface.co/datasets/felipehsilveira/sam3-test.sam3-datasetssam3-door-locksam3-ls-bootstrap-demo
davanstrien/sam3-ls-bootstrap-demo
Bootstrap dataset produced by running facebook/sam3 over a small set of test images and storing the predictions in a Label Studio project for review.
This is a proof-of-concept artifact demonstrating an end-to-end "unlabeled images → bootstrapped dataset" workflow on Hugging Face infrastructure. The predictions in this dataset are SAM3 outputs — not human-reviewed.
Workflow
Images imported into Label Studio project 20 on… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/sam3-ls-bootstrap-demo.imagenes-sam3britain-handbooks-sam3-photos-1500
Object Detection: Photograph Detection using sam3
This dataset contains object detection results (bounding boxes) for photograph detected in images from NationalLibraryOfScotland/Britain-and-UK-Handbooks-Dataset using Meta's SAM3 (Segment Anything Model 3).
Generated using: uv-scripts/sam3 detection script
Detection Statistics
Objects Detected: photograph
Total Detections: 4,500
Images with Detections: 1,500 / 1,500 (100.0%)
Average Detections per Image: 3.00… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/britain-handbooks-sam3-photos-1500.sam3-sam3Sam3Render_InferenceData_v02
