CoolFace
Datasetpublicgated

Faizaniqbal/IndicOCR

Large-scale multilingual OCR and document dataset across 23 Pan-Indic languages and 12 writing systems. 1. Overview IndicOCR (IndicPixel) is a large-scale multilingual Optical Character Recognition (OCR) and document dataset covering the South Asian linguistic landscape. The dataset provides dense document coverage across 23 official and literary languages representing 12 distinct writing systems. Dataset Specifications: Scale & Scope: Over 12… See the full description on the dataset page: https://huggingface.co/datasets/Faizaniqbal/IndicOCR.

sourceHugging Faceapache-2.0updated 7h agoView on Hugging Face
0likes1.2kdownloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.