datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
InsPLAD-workshop-pool
Dataset Card for InsPLAD Workshop Pool
This is a FiftyOne dataset with 1754 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/InsPLAD-workshop-pool")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/InsPLAD-workshop-pool.CVPR_workshop_efficiencyVLMworkshop3uncannyelevationWorkshop-CarDD-Dataset
Dataset Card for cardd_from_hub
This is a FiftyOne dataset with 2816 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Arvind1403/Workshop-CarDD-Dataset")
# Launch the App
session = fo.launch_app(dataset)
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Arvind1403/Workshop-CarDD-Dataset.Workshop-CarDD-Dataset-Subset
Dataset Card for Workshop-CarDD-Dataset
This is a FiftyOne dataset with 500 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Arvind1403/Workshop-CarDD-Dataset-Subset")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/Arvind1403/Workshop-CarDD-Dataset-Subset.marimo-workshop-catalogue
marimo workshop product catalogue
A small, ready-to-search product catalogue for a hands-on marimo workshop, where students build a multimodal
(text and photo) product search engine that runs on a CPU.
File
What it is
catalogue.csv
3,000 products, 150 in each of 20 categories
images.zip
images/<product_id>.jpg, 240×320 JPEG
image_vectors.npy
(3000, 512) float32 CLIP image vectors, one per catalogue row, each of length 1
products.csv
a random 500-row subset used… See the full description on the dataset page: https://huggingface.co/datasets/PS4Research/marimo-workshop-catalogue.spot-telluride-workshop-dataset
Spot Telluride Workshop Dataset
Multimodal sensor data from a Boston Dynamics Spot D02 (Marble backpack compute), extracted from ROS bags for the Telluride Neuromorphic + AI Workshop. The dataset currently covers two locations, each with its own extraction tool and Hub layout root:
Location
Runs
Hub layout root
Extraction tool
Classroom demo
run1
classroom/run1/<modality>/
extract_demo_data.py
School
run1, run2
school/run1/<modality>/, school/run2/<modality>/… See the full description on the dataset page: https://huggingface.co/datasets/lorinachey/spot-telluride-workshop-dataset.cvpr_workshop_sam_prompt_clustersnutrition-workshop-sample
Nutrition Workshop Sample
A small food object detection set for teaching. It is a sample of the
Food Portion Benchmark (FPB),
cut down to 20 dishes so that a beginner can train a working detector on a free
Colab GPU in under ten minutes.
Built for the Nutrition and AI workshop at ISSAI, Nazarbayev University, for
participants with no programming background.
What is in it
Task
object detection, YOLO format
Dishes
20
Photos
1,127
Boxes
1,699… See the full description on the dataset page: https://huggingface.co/datasets/issai/nutrition-workshop-sample.droid-frames-workshop
Dataset Card for droid-frames
This is a FiftyOne dataset with 5172 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("dgural/droid-frames-workshop")
# Launch the App
session = fo.launch_app(dataset)
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/dgural/droid-frames-workshop.cardd_workshop_post_03
Dataset Card for harpreetsahota/cardd_workshop_post_03
This is a FiftyOne dataset with 2816 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/cardd_workshop_post_03")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/cardd_workshop_post_03.cardd_workshop_post_01
Dataset Card for car_dd
This is a FiftyOne dataset with 2816 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/cardd_workshop_post_01")
# Launch the App
session = fo.launch_app(dataset)
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/cardd_workshop_post_01.dataset-public
Auto-Annotation with Expert-Crafted Guidelines: A Study through 3D LiDAR Detection Benchmark
Inspired by the critical bottleneck of data annotation in autonomous driving and recent advancements in foundation models, this dataset introduces a novel evaluation paradigm: Auto-Annotation from Expert-Crafted Guidelines.
Unlike traditional 3D perception benchmarks that rely on massive amounts of annotated 3D point clouds for supervised learning, AutoExpert is designed to evaluate a… See the full description on the dataset page: https://huggingface.co/datasets/autoexpert-cvpr2026-workshop/dataset-public.cardd_workshop_post_threed
Dataset Card for harpreetsahota/cardd_workshop_post_03
This is a FiftyOne dataset with 2816 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/cardd_workshop_post_threed")
# Launch the App
session =… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/cardd_workshop_post_threed.assetscwe-workshop-datasetcardd_workshop_post_02
Dataset Card for harpreetsahota/cardd_workshop_post_01
This is a FiftyOne dataset with 2816 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/cardd_workshop_post_02")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/cardd_workshop_post_02.fe_workshop_assets
