CoolFace
Datasetpublic

wong132/bengali-hindi-number-blindspot

Blind Spots of Frontier Models: Bengali & Hindi Number Word-to-Digit Conversion Summary This dataset documents a critical blind spot in small open-source language models: failure to correctly convert Bengali and Hindi number words into their digit equivalents. Bengali and Hindi share the South Asian number system (hazar/হাজার, lakh/লাখ, crore/কোটি), and all three tested models consistently fail at this fundamental conversion step. IMPORTANT: Arithmetic… See the full description on the dataset page: https://huggingface.co/datasets/wong132/bengali-hindi-number-blindspot.

sourceHugging Facemitupdated 7mo agoView on Hugging Face
0likes17downloads
settings

This repository belongs to wong132 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namebengali-hindi-number-blindspot
visibilitypublic
licencemit
gatedno
ownerwong132
Account settings