datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
KMMLU
KMMLU (Korean-MMLU)
We propose KMMLU, a new Korean benchmark with 35,030 expert-level multiple-choice questions across 45 subjects ranging from humanities to STEM.
Unlike previous Korean benchmarks that are translated from existing English benchmarks, KMMLU is collected from original Korean exams, capturing linguistic and cultural aspects of the Korean language.
We test 26 publically available and proprietary LLMs, identifying significant room for improvement.
The best publicly… See the full description on the dataset page: https://huggingface.co/datasets/HAERAE-HUB/KMMLU.KMMLU-Summarized-Chain_of_Thought
Dataset Card for Condensed Chain-of-Thought KMMLU Dataset
This dataset card provides detailed information about the condensed KMMLU dataset. The dataset has been summarized using Upstage's LLM: Solar-Pro to condense the original KMMLU training and development data while preserving its quality and usability. Additionally, a new column, 'chain_of_thought', has been introduced to align with the reasoning approach outlined in the paper "Chain-of-Thought Prompting Elicits Reasoning in… See the full description on the dataset page: https://huggingface.co/datasets/SabaPivot/KMMLU-Summarized-Chain_of_Thought.KMMLU-HARD
KMMLU (Korean-MMLU)
We propose KMMLU, a new Korean benchmark with 35,030 expert-level multiple-choice questions across 45 subjects ranging from humanities to STEM.
Unlike previous Korean benchmarks that are translated from existing English benchmarks, KMMLU is collected from original Korean exams, capturing linguistic and cultural aspects of the Korean language.
We test 26 publically available and proprietary LLMs, identifying significant room for improvement.
The best publicly… See the full description on the dataset page: https://huggingface.co/datasets/HAERAE-HUB/KMMLU-HARD.KMMLU-Pro
KMMLU-Pro
📄 Paper | 🖥️ Code
We introduce KMMLU-PRO, a challenging new benchmark comprising 2,822 problems from the official exams for Korean National Professional Licensure (KNPL), representing highly specialized professions in Korea. To bridge the gap between LLM performance and real-world applicability, our evaluation closely mirrors the official certification criteria including reporting the number of professional licenses each LLM could pass according to these standards. All… See the full description on the dataset page: https://huggingface.co/datasets/LGAI-EXAONE/KMMLU-Pro.KMMLU-Redux
KMMLU-Redux
📄 Paper
We introduce KMMLU-Redux, a reconstructed version of the existing KMMLU, comprising 2,587 problems from Korean National Technical Qualification (KNTQ) exams. We identified several critical issues in the KMMLU, including leaked answers, lack of clarity, ill-posed questions, notation errors, and contamination risks. To address these problems, we conducted rigorous manual examination and decontaminated the dataset from the pre-training corpus to prevent potential… See the full description on the dataset page: https://huggingface.co/datasets/LGAI-EXAONE/KMMLU-Redux.KMMLU-en
KMMLU-en (English-Translated KMMLU)
This dataset is the English-translated version of the original KMMLU (Korean-MMLU) dataset.
Description
The original KMMLU is a challenging benchmark designed to measure massive multitask language understanding in Korean. It consists of 35,030 expert-level multiple-choice questions across 45 diverse subjects, collected from original Korean exams to capture unique linguistic and cultural nuances.
KMMLU-en was created to enable evaluation… See the full description on the dataset page: https://huggingface.co/datasets/bzantium/KMMLU-en.kmmlu-pro-doctorkmmlukmmlu-formatted
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
KMMLU 통합본
base dataset: HAERAE-HUB/KMMLU
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/richard-park/kmmlu-formatted.kmmlu-conversation-sampleKmmlu 데이터를 이용해서 대화 데이터셋 샘픔을 생성하였습니다.
멀티턴 데이터셋으로 학습용도로 만들어졌습니다.
kmmlu-flatKMMLU-DPOKMMLU_Redux_MCKMMLU-EASYScholarBench-as-KMMLUkmmlu_testkmmlu90_testkmmlu_subset100kmmlu-train-reasoningko-syn-kmmluko-sft-v1024-without-kmmlukmmlu_trainset
