RaspizdAI/QUAD-Bench
QUAD-Bench: A Lightweight AI Reasoning & Acuity Benchmark QUAD-Bench is a compact, multiple-choice benchmark dataset designed for quick evaluation of Large Language Models (LLMs). The dataset contains 100 questions evenly distributed across 4 fundamental capabilities, requiring the model to select exactly one correct answer option (A, B, C, or D). 📊 Dataset Structure The benchmark consists of 100 questions divided into 4 categories (25 questions each):… See the full description on the dataset page: https://huggingface.co/datasets/RaspizdAI/QUAD-Bench.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face