CoolFace
Datasetpublic

meharuhanzz/OCR-Bench1000-Punjabi

OCR-Bench1000-Punjabi 1000 synthetic printed-text line images with ground-truth transcriptions, sampled from a larger locally-held Punjabi OCR training corpus. This is a benchmark/sample release, not the full training set. Data fields Field Description file_name relative path to the image (images/...) text ground-truth transcription category punjabi_only / english_only / mixed / numeric_and_symbols length_bucket short / medium / long, by character… See the full description on the dataset page: https://huggingface.co/datasets/meharuhanzz/OCR-Bench1000-Punjabi.

sourceHugging Facemitupdated 13d agoView on Hugging Face
0likes157downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
meharuhanzz/OCR-Bench1000-Punjabi · CoolFace