CoolFace
Datasetpublic

saeid1999/fa-en-ar-handwritten-ocr-v1

Multi-script Synthetic Handwritten OCR — fa / ar / en A large, clean, augmentation-rich synthetic handwriting dataset for training and benchmarking OCR / HTR models on Persian (fa), Arabic (ar) and English (en). Every line image ships with an exact Unicode transcription plus rich provenance metadata (writer style, font, ink, script direction, digit system). Page-level PAGE-XML and COCO ground truth support layout-aware training and evaluation out of the box. 1,000 rendered… See the full description on the dataset page: https://huggingface.co/datasets/saeid1999/fa-en-ar-handwritten-ocr-v1.

sourceHugging Facecc-by-4.0updated 10d agoView on Hugging Face
0likes415downloads
diriam/
fileaugmented.zip107.5 MBdownload
filewords.zip750.3 MBdownload

saeid1999/fa-en-ar-handwritten-ocr-v1 · main · files are served by the source, never re-hosted here