CoolFace
Datasetpublic

Werea-co/werea-tr-doc-ocr-synthetic

Werea Turkish Enterprise Documents 📄🇹🇷 Türkçe kurumsal belge OCR eğitimi için tamamı sentetik sayfa görüntüleri ve birebir eşleşen markdown ground-truth metinleri. Werea tarafından Werea-DocOCR modellerinin eğitimi için üretilmiştir. Belge türleri Tür Train Test İçerik Genel vekaletname 500 50 Noter başlığı, taraflar, yetki maddeleri, noter şerhi DASK poliçesi 500 50 Poliçe/sigortalı/bina bilgileri, prim tablosu e-Arşiv fatura 500 50 Satıcı/alıcı… See the full description on the dataset page: https://huggingface.co/datasets/Werea-co/werea-tr-doc-ocr-synthetic.

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
0likes375downloads
10 commits on main
1a884f61mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 9)

GoktugD
a8ec3501mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 8)

GoktugD
83fe6851mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 7)

GoktugD
a53b2ad1mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 6)

GoktugD
3bbee8f1mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 5)

GoktugD
732c0521mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 4)

GoktugD
2db71611mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 3)

GoktugD
b485f3f1mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts) (part 2)

GoktugD
95983da1mo ago

Add synthetic Turkish enterprise document OCR dataset (3000 train + 300 test + generator/train scripts)

GoktugD
5deb3451mo ago

initial commit

GoktugD