CoolFace
Datasetpublic

ysober/zh_spec_eval

中文专项评测集 本评测集共包含 512 条样本,分为 4 个工作负载(workload),每个工作负载包含 128 条样本。数据文件位于当前目录。 数据概览 Workload 来源数据集 数据划分 采样方法 Prompt 长度中位数(token) zh_ceval ceval/ceval-exam val 汇总全部 52 个学科的样本,使用随机种子 0 打乱后取前 128 条,以兼顾学科覆盖的均衡性 96 zh_gaokao_math hails/agieval-gaokao-mathqa test(351 条) 使用随机种子 0 打乱后取前 128 条 142 zh_simpleqa OpenStellarTeam/Chinese-SimpleQA train(3,000 条) 使用随机种子 0 打乱后取前 128 条 30 zh_alpaca_gpt4 llm-wizard/alpaca-gpt4-data-zh train(48,818 条) 过滤掉 instruction 少于… See the full description on the dataset page: https://huggingface.co/datasets/ysober/zh_spec_eval.

sourceHugging Faceupdated 1mo agoView on Hugging Face
0likes81downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ysober/zh_spec_eval · CoolFace