CoolFace
Datasetpublic

HumynLabs/Italian_Documents_Dataset_PDF

Italian Documents Dataset (PDF) This dataset contains a curated collection of Italian-language documents in PDF format. It includes books, academic publications, reports, government documents, and news articles written in Italian. The dataset supports AI research in OCR, multilingual document understanding, and text recognition for Romance languages. Contact For queries or collaborations related to this dataset, contact: anoushka@kgen.io… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/Italian_Documents_Dataset_PDF.

sourceHugging Facecc-by-4.0updated 11mo agoView on Hugging Face
0likes724downloads
settings

This repository belongs to HumynLabs on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameItalian_Documents_Dataset_PDF
visibilitypublic
licencecc-by-4.0
gatedno
ownerHumynLabs
Account settings
HumynLabs/Italian_Documents_Dataset_PDF · CoolFace