tom-jerry-123/Physical-AI-AV-DE
PhysicalAI-AV-SFT Supervised fine-tuning (SFT) dataset for an autonomous-vehicle vision-language waypoint-prediction model. Contains 324,105 samples from 150 000 driving scenes (18 seconds per scene, sampled at anchor times 2s..16s) recorded in the United States. Format WebDataset — 10 uncompressed .tar shards, each containing pairs of files per sample: Entry Description {key}.jpg Front-facing wide-angle camera frame (JPEG quality 95, 640 × 360 px)… See the full description on the dataset page: https://huggingface.co/datasets/tom-jerry-123/Physical-AI-AV-DE.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face