CoolFace
Datasetpublic

ved1245/synthetic-manuscript-dataset

Synthetic Manuscript Dataset Synthetic historical manuscript folios generated using an automated Python pipeline. Scripts The dataset contains three script configurations: Devanagari Modi Sharada Each script contains 100 synthetic manuscript folios. Dataset Splits Split Samples per Script Train 85 Validation 10 Test 5 Total 100 Across all three scripts, the dataset contains: 300 manuscript images 300 corresponding Markdown… See the full description on the dataset page: https://huggingface.co/datasets/ved1245/synthetic-manuscript-dataset.

sourceHugging Faceupdated 15d agoView on Hugging Face
1likes776downloads
settings

This repository belongs to ved1245 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namesynthetic-manuscript-dataset
visibilitypublic
licencenot set
gatedno
ownerved1245
Account settings
ved1245/synthetic-manuscript-dataset · CoolFace