CoolFace
Datasetpublic

opencsg/autohub-benchmark

autohub-benchmark This project designs common use scenarios for web-based code, model, and dataset hosting platforms, and provides corresponding prompts and ground truth. These resources can be used to evaluate the localization performance of visual language models (VLMs) in specialized scenarios. Model Hosting Platform GUI Inference Model Platform Accuracy (%) Error (%) Invalid (%) Completion Rate (%) AriaUI Huggingface 70.8 12.5 6.7 100.0… See the full description on the dataset page: https://huggingface.co/datasets/opencsg/autohub-benchmark.

sourceHugging Faceupdated 2y agoView on Hugging Face
1likes158downloads
3_0_datasets.png4 linesDownload Raw Back to ModelScope
1version https://git-lfs.github.com/spec/v12oid sha256:368d4d22475457404b28980e5333a385f6eb0b9e411bb3154e72152e0a07096d3size 4100454