datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v12-Prose-DS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v12-Prose-DS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details.sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details
Dataset Card for Evaluation run of sometimesanotion/Qwen-2.5-14B-Virmarckeoso
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-2.5-14B-Virmarckeoso
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details.sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details.sometimesanotion__Qwentinuum-14B-v7-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v7
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v7
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v7-details.sometimesanotion__Qwen-14B-ProseStock-v4-details
Dataset Card for Evaluation run of sometimesanotion/Qwen-14B-ProseStock-v4
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-14B-ProseStock-v4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-14B-ProseStock-v4-details.sometimesanotion__LamarckInfusion-14B-v2-details
Dataset Card for Evaluation run of sometimesanotion/LamarckInfusion-14B-v2
Dataset automatically created during the evaluation run of model sometimesanotion/LamarckInfusion-14B-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__LamarckInfusion-14B-v2-details.sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details.sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v13-Prose-DS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v13-Prose-DS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details.some-rp-v2This is a filtered subset of lemonilia/Elliquiy-Role-Playing-Forums_2023-04 that's been converted into a typical turn-based conversation format (the added conversations column). These are all 2-user chats, and the first user was assigned to human and the second to gpt.
someone
someone* 订单筛选结果
来源:/home/GRQ/lw_9/ordermsg_logs_20260910。每单两份文件:<日期>_<order_id>.llm_record.jsonl(该单全部模型调用,按时间排序)、<日期>_<order_id>.app.log(该单全部运行日志行)。共 58 单。
订单
日期
产品
回复轮
detect 次
跨度(分)
日志行
20260910_someone1125_test1_2694
20260910
psyche-link
13
8
7.2
392
20260910_someone115_test1_6967
20260910
psyche-link
10
7
6.4
319
20260910_someone1186_test1_1252
20260910
psyche-link
12
7
7.3
356
20260910_someone1286_test1_562
20260910
psyche-link
8
0
6.1
151… See the full description on the dataset page: https://huggingface.co/datasets/HYGGEhygge/someone.sometimesanotion__IF-reasoning-experiment-80-details
Dataset Card for Evaluation run of sometimesanotion/IF-reasoning-experiment-80
Dataset automatically created during the evaluation run of model sometimesanotion/IF-reasoning-experiment-80
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__IF-reasoning-experiment-80-details.sometimesanotion__Lamarck-14B-v0.7-Fusion-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.7-Fusion
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.7-Fusion
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.7-Fusion-details.sometimesanotion__Lamarck-14B-v0.3-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.3
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.3-details.sometimesanotion__lamarck-14b-prose-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-prose-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-prose-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-prose-model_stock-details.sometimesanotion__lamarck-14b-reason-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-reason-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-reason-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-reason-model_stock-details.DreadPoor__Something-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Something-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Something-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Something-8B-Model_Stock-details.sometimesanotion__Lamarck-14B-v0.4-Qwenvergence-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.4-Qwenvergence
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.4-Qwenvergence
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.4-Qwenvergence-details.sometimesanotion__Qwenvergence-14B-v15-Prose-MS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v15-Prose-MS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v15-Prose-MS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v15-Prose-MS-details.sometimesanotion__Qwentinuum-14B-v2-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v2
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v2-details.sometimesanotion__Lamarck-14B-v0.6-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.6
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.6
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.6-details.sometimesanotion__Qwenvergence-14B-v3-Prose-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v3-Prose
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v3-Prose
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v3-Prose-details.sometimesanotion__Qwenvergence-14B-v6-Prose-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v6-Prose-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v6-Prose-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v6-Prose-model_stock-details.sometimesanotion__Qwenvergence-14B-v6-Prose-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v6-Prose
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v6-Prose
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v6-Prose-details.sometimesanotion__Qwenvergence-14B-v3-Reason-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v3-Reason
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v3-Reason
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v3-Reason-details.sometimesanotion__Qwentinuum-14B-v5-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v5
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v5
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v5-details.sometimesanotion__Qwentinuum-14B-v6-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v6
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v6
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v6-details.sometimesanotion__Qwentinuum-14B-v6-Prose-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v6-Prose
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v6-Prose
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v6-Prose-details.sometimesanotion__IF-reasoning-experiment-40-details
Dataset Card for Evaluation run of sometimesanotion/IF-reasoning-experiment-40
Dataset automatically created during the evaluation run of model sometimesanotion/IF-reasoning-experiment-40
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__IF-reasoning-experiment-40-details.sometimesanotion__Lamarck-14B-v0.6-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.6-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.6-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.6-model_stock-details.sumink__somer-details
Dataset Card for Evaluation run of sumink/somer
Dataset automatically created during the evaluation run of model sumink/somer
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sumink__somer-details.
