cis-lmu/GlotOCR-bench
GlotOCR-bench GlotOCR-bench is a dataset of 16375 images covering 158 writing systems (+2000 languages), designed to evaluate the fundamental OCR capabilities required to support diverse writing systems and languages. Quick links: π Paper π οΈ Code π Results π Leaderboard License This dataset is released under the GlotOCR Open Evaluation License v1.0 (see LICENSE file for full terms). The GlotOCR-bench metadata is licensed under CC0-1.0. The texts used toβ¦ See the full description on the dataset page: https://huggingface.co/datasets/cis-lmu/GlotOCR-bench.
645
No card is published for this repository, or it could not be fetched from Hugging Face right now.
