datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DocStruct4MDocReason25Kdoc_vla_cache_train
NAVSIM navtrain metric cache
NAVSIM navtrain split 的 metric cache,用于 PDM score 计算(AutoVLA 等做 RL/GRPO 训练时的 reward,
或跑 PDMS 评测)。纯 CPU 产物,与模型无关,一次生成可永久复用。
场景数
103,288(train 101,288 + val 2,000)
大小
35 GB(解包后)
生成
navsim/planning/script/run_metric_caching.py,train_test_split=navtrain
覆盖率
对 navtrain 的 101,288 个训练样本 100% 覆盖,缺 0
每个 metric_cache.pkl(lzma 压缩,约 450 KB)含 PDMScorer 判分所需的全部内容:
ego_state、trajectory(PDM 参考轨迹)、observation(各时刻 agent 占用)、… See the full description on the dataset page: https://huggingface.co/datasets/PhoenixHu/doc_vla_cache_train.docmatixdocker_imageswormhole-docsdoc-audio-11
[doc] audio dataset 11
This dataset contains two tar files that contain pairs of samples with one audio file and one JSON file.
lhy_docvqa_datavoyager_docker_imagedocvqa-privacy-dataCram_school_doclayoutAgent_document_1Do_CLIP
