datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
efficient-vla-extracts
efficient-vla-extracts
Parser-derived artifacts for the efficient-vla-wiki project.
This dataset is the data layer companion to the main repository:
GitHub: guanweifan/efficient-vla-wiki
Dataset repo: efficient-vla-extracts
What is inside
The staged upload keeps the local directory structure:
extracts/
├── meta/
│ ├── extract_build_index.jsonl
│ └── extract_build_status.json
└── parses/
└── <paper_id>/
Each paper directory may contain:
pdftotext.txt… See the full description on the dataset page: https://huggingface.co/datasets/guanweifan/efficient-vla-extracts.EndoBench-Extended
EndoBench
🍎 Homepage|💻 GitHub|🤗 Dataset|📖 Paper
This repository is the official implementation of the paper EndoBench: A Comprehensive Evaluation of Multi-Modal Large Language Models for Endoscopy Analysis.
🚀 News
[21/10/2025] We release a new open-set challenging VQA benchmark EndoBench-extended.
[19/09/2025] 🎉🎉Our EndoBench was accepted by NeurIPS'25 D&B Track!!!
🏥 EndoBench-Extended
This EndoBench-extendeddataset is the extended version of… See the full description on the dataset page: https://huggingface.co/datasets/Saint-lsy/EndoBench-Extended.
