CoolFace
Datasetpublic

ScaleAI/VisualToolBench

VisToolBench Dataset A benchmark dataset for evaluating vision-language models on tool-use tasks. Dataset Statistics Total samples: 1204 Single-turn: 603 Multi-turn: 601 Schema Column Type Description id string Unique task identifier turncase string Either "single-turn" or "multi-turn" num_turns int Number of conversation turns (1 for single-turn) prompt_category string Task category (e.g., "medical", "scientific", "general")… See the full description on the dataset page: https://huggingface.co/datasets/ScaleAI/VisualToolBench.

sourceHugging Faceupdated 9mo agoView on Hugging Face
7likes614downloads
14 commits on main
28de4fb9mo ago

Upload VisualToolBench test split

utkarsh4430
4b085559mo ago

Upload folder using huggingface_hub

utkarsh4430
3f806409mo ago

Upload folder using huggingface_hub

utkarsh4430
0f719889mo ago

Upload folder using huggingface_hub

utkarsh4430
040d8d49mo ago

Delete manifest.json

yunzhong-scale
6b747f09mo ago

Delete vistoolbench_1204.parquet

yunzhong-scale
407044d9mo ago

Update README.md

yunzhong-scale
acb511b9mo ago

Update corrected data + README metadata

yunzhong-scale
8fed9dd9mo ago

Upload folder using huggingface_hub

utkarsh4430
501c32811mo ago

Update README.md

yunzhong-scale
c1837e411mo ago

Update README.md

yunzhong-scale
0308b8d11mo ago

Create README.md

yunzhong-scale
559a0ed11mo ago

Initial upload

yunzhong-scale
5e66d3f11mo ago

initial commit

yunzhong-scale