CoolFace
Datasetpublic

Nasser4963/jamo-gold

JAMO Gold — Human-Validated Hangul Rendering Judgments A small, carefully labelled dataset for a specific purpose: testing whether an OCR engine can tell that a glyph is not a real character. It is not a Hangul OCR training set, and not a model leaderboard. Why this exists Benchmarks that measure visual text rendering score generated images with OCR. But OCR engines are closed-set classifiers — given an image, they must return some character from their… See the full description on the dataset page: https://huggingface.co/datasets/Nasser4963/jamo-gold.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
1likes80downloads
7 commits on main
f2afeb71mo ago

Regenerate PDF from corrected TECHNICAL_NOTE.ko.md (Reproducibility section)

Nasser4963
0d9fc4f1mo ago

Regenerate PDF from corrected TECHNICAL_NOTE.md (Reproducibility section)

Nasser4963
7a28b6d1mo ago

Fix Reproducibility section (KO): point to github.com/Nasser-Lim/jamo-bench

Nasser4963
cf2565e1mo ago

Fix Reproducibility section: point to github.com/Nasser-Lim/jamo-bench

Nasser4963
5d020e62mo ago

Add authorship/AI tool-use disclosure (note S10); refresh SEO metadata and quickstart

Nasser4963
692164b2mo ago

Initial release: JAMO Gold dataset + technical note

Nasser4963
2bf0bb22mo ago

initial commit

Nasser4963