CoolFace
Datasetpublicgated

cis-lmu/GlotOCR-bench

GlotOCR-bench GlotOCR-bench is a dataset of 16375 images covering 158 writing systems (+2000 languages), designed to evaluate the fundamental OCR capabilities required to support diverse writing systems and languages. Quick links: ๐Ÿ“ƒ Paper ๐Ÿ› ๏ธ Code ๐Ÿ“ˆ Results ๐Ÿ† Leaderboard License This dataset is released under the GlotOCR Open Evaluation License v1.0 (see LICENSE file for full terms). The GlotOCR-bench metadata is licensed under CC0-1.0. The texts used toโ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/cis-lmu/GlotOCR-bench.

sourceHugging Faceotherupdated 6mo agoView on Hugging Face
6likes40downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone elseโ€™s repository from here would need an authorised integration and the account holderโ€™s consent, so the link goes to the source instead.

Open discussions on Hugging Face
cis-lmu/GlotOCR-bench ยท CoolFace