arobin79/bangla-ocr-validation_data_printed
Bangla OCR Validation Dataset (Printed + Scanned) 📌 Description This dataset is a Bangla OCR validation dataset containing a mix of printed document images and their corresponding text annotations. It is designed to evaluate OCR and vision-language models on both clean digital text and scanned document images. 📊 Dataset Composition 1507 line-level images with text annotations 50 full-page document images with text Data includes: Printed/typed… See the full description on the dataset page: https://huggingface.co/datasets/arobin79/bangla-ocr-validation_data_printed.
This repository belongs to arobin79 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
