CoolFace
Datasetpublic

opencsg/autohub-benchmark

autohub-benchmark This project designs common use scenarios for web-based code, model, and dataset hosting platforms, and provides corresponding prompts and ground truth. These resources can be used to evaluate the localization performance of visual language models (VLMs) in specialized scenarios. Model Hosting Platform GUI Inference Model Platform Accuracy (%) Error (%) Invalid (%) Completion Rate (%) AriaUI Huggingface 70.8 12.5 6.7 100.0… See the full description on the dataset page: https://huggingface.co/datasets/opencsg/autohub-benchmark.

sourceHugging Faceupdated 2y agoView on Hugging Face
1likes158downloads
state.json13 linesDownload Raw Back to data
1{2  "_data_files": [3    {4      "filename": "data-00000-of-00001.arrow"5    }6  ],7  "_fingerprint": "d8c807064434a661",8  "_format_columns": null,9  "_format_kwargs": {},10  "_format_type": null,11  "_output_all_columns": false,12  "_split": null13}