ysober/zh_spec_eval
中文专项评测集 本评测集共包含 512 条样本,分为 4 个工作负载(workload),每个工作负载包含 128 条样本。数据文件位于当前目录。 数据概览 Workload 来源数据集 数据划分 采样方法 Prompt 长度中位数(token) zh_ceval ceval/ceval-exam val 汇总全部 52 个学科的样本,使用随机种子 0 打乱后取前 128 条,以兼顾学科覆盖的均衡性 96 zh_gaokao_math hails/agieval-gaokao-mathqa test(351 条) 使用随机种子 0 打乱后取前 128 条 142 zh_simpleqa OpenStellarTeam/Chinese-SimpleQA train(3,000 条) 使用随机种子 0 打乱后取前 128 条 30 zh_alpaca_gpt4 llm-wizard/alpaca-gpt4-data-zh train(48,818 条) 过滤掉 instruction 少于… See the full description on the dataset page: https://huggingface.co/datasets/ysober/zh_spec_eval.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face