CoolFace
Datasetpublic

ilsilfverskiold/ocr-benchmark

OCR Benchmark — Documents The 93 document images and ground truth used by the ocr-benchmark harness. The benchmark code, the reference run results, and the full methodology live in the GitHub repo — this dataset is the document corpus only. Structure One train split, 93 rows, one row per document: Column Type Description image Image The document page (PNG/JPG) stem string Filename stem (e.g. invoice_000) tier string Difficulty: easy, medium, or hard… See the full description on the dataset page: https://huggingface.co/datasets/ilsilfverskiold/ocr-benchmark.

sourceHugging Facemitupdated 2mo agoView on Hugging Face
0likes489downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ilsilfverskiold/ocr-benchmark · CoolFace