CoolFace
Datasetpublic

ved1245/synthetic-manuscript-dataset

Synthetic Manuscript Dataset Synthetic historical manuscript folios generated using an automated Python pipeline. Scripts The dataset contains three script configurations: Devanagari Modi Sharada Each script contains 100 synthetic manuscript folios. Dataset Splits Split Samples per Script Train 85 Validation 10 Test 5 Total 100 Across all three scripts, the dataset contains: 300 manuscript images 300 corresponding Markdown… See the full description on the dataset page: https://huggingface.co/datasets/ved1245/synthetic-manuscript-dataset.

sourceHugging Faceupdated 15d agoView on Hugging Face
1likes778downloads

ved1245/synthetic-manuscript-dataset · main · files are served by the source, never re-hosted here