wong132/bengali-hindi-number-blindspot
Blind Spots of Frontier Models: Bengali & Hindi Number Word-to-Digit Conversion Summary This dataset documents a critical blind spot in small open-source language models: failure to correctly convert Bengali and Hindi number words into their digit equivalents. Bengali and Hindi share the South Asian number system (hazar/হাজার, lakh/লাখ, crore/কোটি), and all three tested models consistently fail at this fundamental conversion step. IMPORTANT: Arithmetic… See the full description on the dataset page: https://huggingface.co/datasets/wong132/bengali-hindi-number-blindspot.
Upload README.md with huggingface_hub
Upload README.md with huggingface_hub
Delete Untitled10.ipynb with huggingface_hub
Upload README.md with huggingface_hub
Upload Untitled10.ipynb with huggingface_hub
Upload tiny-aya-global.csv with huggingface_hub
Upload tiny-aya-fire.csv with huggingface_hub
Upload qwen3.5-4b.csv with huggingface_hub
Upload README.md with huggingface_hub
initial commit
