CoolFace
Datasetpublicgated

naist-nlp/SinhalaMMLU

SinhalaMMLU We introduce SinhalaMMLU, the first multiple-choice question answering benchmark designed specifically for Sinhala, a low-resource language.The dataset contains over 7,000 questions spanning secondary to collegiate education levels, aligned with the Sri Lankan national curriculum.It covers six domains and 30 subjects, encompassing both general academic topics and culturally grounded knowledge. We evaluated 26 large language models (LLMs) on SinhalaMMLU and observed… See the full description on the dataset page: https://huggingface.co/datasets/naist-nlp/SinhalaMMLU.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes21downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
naist-nlp/SinhalaMMLU · CoolFace