Moealsarraj/arabic-bench-dataset
Arabic Bench Dataset A curated evaluation dataset for benchmarking AI models on Arabic language tasks. Overview 50 test cases across 8 categories Each case includes a prompt, gold-standard reference answer, and a deliberately imperfect AI response Covers Modern Standard Arabic (MSA) and multiple Arabic dialects Designed for evaluating: translation, summarization, Q&A, creative writing, grammar, dialect understanding, legal/formal, and medical/scientific tasks… See the full description on the dataset page: https://huggingface.co/datasets/Moealsarraj/arabic-bench-dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face