datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
physical-ai-bench-generation
Physical AI Bench - Generation
Paper | Code
Dataset Description
The PAI-Bench is a benchmark to measure the progress of world models quantitatively.
The predict task contains a list of 1044 samples of text prompts, conditioning images, and qa pairs, covering Physical AI target domains including autonomous vehicle (AV) driving, robotics, industry (smart space), physics, human, and common sense. All the questions are binary questions, and the answer is either Yes or No. Our… See the full description on the dataset page: https://huggingface.co/datasets/shi-labs/physical-ai-bench-generation.PhysicalAI-ADE-US
PhysicalAI-US-ADE
Dataset Summary
PhysicalAI-US-ADE contains per-sample evaluation outputs for autonomous driving waypoint prediction on the US subset of the PhysicalAI NVIDIA dataset.
This dataset stores inference-time predictions and evaluation statistics for models evaluated on the dataset, organized by model name at the top level. Each model directory contains sample-level records for that model’s predictions against ground truth.
The current release includes… See the full description on the dataset page: https://huggingface.co/datasets/mjf-su/PhysicalAI-ADE-US.
