HAERAE-HUB/HAE_RAE_BENCH_1.1
The HAE_RAE_BENCH 1.1 is an ongoing project to develop a suite of evaluation tasks designed to test the understanding of models regarding Korean cultural and contextual nuances. Currently, it comprises 13 distinct tasks, with a total of 4900 instances. Please note that although this repository contains datasets from the original HAE-RAE BENCH paper, the contents are not completely identical. Specifically, the reading comprehension subset from the original version has been removed due to… See the full description on the dataset page: https://huggingface.co/datasets/HAERAE-HUB/HAE_RAE_BENCH_1.1.
The HAERAEBENCH 1.1 is an ongoing project to develop a suite of evaluation tasks designed to test the understanding of models regarding Korean cultural and contextual nuances. Currently, it comprises 13 distinct tasks, with a total of 4900 instances.
Please note that although this repository contains datasets from the original HAE-RAE BENCH paper, the contents are not completely identical. Specifically, the reading comprehension subset from the original version has been removed due to copyright constraints. In its place, an updated reading comprehension subset has been introduced, sourced from the CSAT, the Korean university entrance examination. To replicate the studies from the paper, please see code.
Dataset Overview
Point of Contact
For any questions contact us via the following email:)
spthsrbwls123@yonsei.ac.kr