CoolFace
Datasetpublic

typhoon-ai/ThaiOCRBench

ThaiOCRBench: A Task-Diverse Benchmark for Vision-Language Understanding in Thai ThaiOCRBench is the first comprehensive benchmark for evaluating vision-language models (VLMs) on Thai text-rich visual understanding tasks.Inspired by OCRBench v2, it contains 2,808 human-annotated samples across 13 diverse tasks, including table parsing, chart understanding, full-page OCR, key information extraction, and visual question answering. The benchmark enables standardized zero-shot… See the full description on the dataset page: https://huggingface.co/datasets/typhoon-ai/ThaiOCRBench.

sourceHugging Facecc-by-sa-4.0updated 10mo agoView on Hugging Face
7likes795downloads
../
filetest-00000-of-00004.parquet939.2 MBdownload
filetest-00001-of-00004.parquet351.0 MBdownload
filetest-00002-of-00004.parquet185.6 MBdownload
filetest-00003-of-00004.parquet338.1 MBdownload

typhoon-ai/ThaiOCRBench · main · files are served by the source, never re-hosted here