prithivMLmods/Openpdf-MultiReceipt-1K
Openpdf-MultiReceipt-1K Openpdf-MultiReceipt-1K is a dataset consisting of over 1,000 receipt documents in PDF format. This dataset is designed for use in image-to-text and document understanding tasks, particularly Optical Character Recognition (OCR), receipt parsing, and layout analysis. Notes No text annotations or metadata are provided — only the raw PDFs. Ideal for tasks requiring raw document inputs like PDF-to-Text pipelines. Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Openpdf-MultiReceipt-1K.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face