CoolFace
Datasetpublic

davanstrien/encyclopaedia_britannica_illustrated-lighton-ocr-32k-test

Document OCR using LightOnOCR-0.9B-32k-1025 This dataset contains OCR results from images in NationalLibraryOfScotland/Britain-and-UK-Handbooks-Dataset using LightOnOCR, a fast and compact 1B OCR model. Processing Details Source Dataset: NationalLibraryOfScotland/Britain-and-UK-Handbooks-Dataset Model: lightonai/LightOnOCR-0.9B-32k-1025 Vocabulary Size: 32k tokens Number of Samples: 100 Processing Time: 3.7 min Processing Date: 2025-10-23 17:51 UTC… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/encyclopaedia_britannica_illustrated-lighton-ocr-32k-test.

sourceHugging Faceupdated 11mo agoView on Hugging Face
2likes17downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
davanstrien/encyclopaedia_britannica_illustrated-lighton-ocr-32k-test · CoolFace