CoolFace
Datasetpublic

premmm/nepali-bench

NepaliBench ๐Ÿ”๏ธ A rigorous evaluation benchmark for Nepali language models. Why this exists There is no standard, publicly reproducible benchmark for evaluating Nepali LLMs. This dataset was created after a systematic evaluation of himalaya-ai's NanochatGPT and Gemma fine-tune revealed that models claiming Nepali capability had no shared evaluation standard to measure against. Dataset 100 carefully curated evaluation examples across 8 categories:โ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/premmm/nepali-bench.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
1likes37downloads
3 commits on main
78e02773mo ago

Upload README.md with huggingface_hub

premmm
fd80aba3mo ago

Upload nepali_bench.json with huggingface_hub

premmm
d26cc393mo ago

initial commit

premmm