CoolFace
Datasetpublic

Sigurdur/icelandic-ocr-benchmark

Dataset Card for Icelandic OCR Benchmark Dataset Details Dataset Description Icelandic OCR Benchmark is a ground-truth dataset for evaluating OCR accuracy on Icelandic-language documents. It consists of manually transcribed page images with matching layout annotations (text regions, line polygons, baselines) in both ALTO and PAGE XML. Curated by: Sigurdur Haukur Birgisson Language(s): Icelandic (is) License: CC BY-SA 4.0 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Sigurdur/icelandic-ocr-benchmark.

sourceHugging Facecc-by-sa-4.0updated 10d agoView on Hugging Face
1likes87downloads
settings

This repository belongs to Sigurdur on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameicelandic-ocr-benchmark
visibilitypublic
licencecc-by-sa-4.0
gatedno
ownerSigurdur
Account settings
Sigurdur/icelandic-ocr-benchmark · CoolFace