datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MULocBenchThis repository hosts MULocBench, a comprehensive dataset.
MULocBench addresses limitations in existing benchmarks by focusing on accurate project localization (e.g., files and functions) for issue resolution, which is a critical first step in software maintenance. It comprises 1,100 issues from 46 popular GitHub Python projects, offering greater diversity in issue types, root causes, location scopes, and file types compared to prior datasets. This dataset provides a more realistic testbed for… See the full description on the dataset page: https://huggingface.co/datasets/somethingone/MULocBench.something-something-v2Something-Something-v2smoltalk-binidxsometimesanotion__Qwenvergence-14B-v12-Prose-DS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v12-Prose-DS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v12-Prose-DS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details.something-something-v2sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details
Dataset Card for Evaluation run of sometimesanotion/Qwen-2.5-14B-Virmarckeoso
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-2.5-14B-Virmarckeoso
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details.sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details.sometimesanotion__Qwentinuum-14B-v7-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v7
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v7
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v7-details.sometimesanotion__Qwen-14B-ProseStock-v4-details
Dataset Card for Evaluation run of sometimesanotion/Qwen-14B-ProseStock-v4
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-14B-ProseStock-v4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-14B-ProseStock-v4-details.sometimesanotion__LamarckInfusion-14B-v2-details
Dataset Card for Evaluation run of sometimesanotion/LamarckInfusion-14B-v2
Dataset automatically created during the evaluation run of model sometimesanotion/LamarckInfusion-14B-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__LamarckInfusion-14B-v2-details.some
FinEE Dataset
Dataset Description
A comprehensive dataset for training financial entity extraction models on Indian banking messages. Contains 152,000+ samples covering SMS, emails, and transaction notifications from major Indian banks.
Languages
English (en) - 86%
Hindi (hi) - 3%
Tamil (ta) - 3%
Telugu (te) - 3%
Bengali (bn) - 3%
Kannada (kn) - 2%
Supported Transaction Types
UPI payments (PhonePe, GPay, Paytm)… See the full description on the dataset page: https://huggingface.co/datasets/Siddhu077/some.sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details.sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v13-Prose-DS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v13-Prose-DS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details.Tieba-SomeInteresting
中文数据集(基于百度贴吧)
简介
本数据集数据来自于2025年3月5日百度贴吧
数据集使用了“孙笑川吧”、“弱智吧”、“中国人口吧”和“航空母舰吧”
数据集问题来自于发帖的标题,答案来自于最热门的回复
思考链来自于 DeepSeek-v3 生成
目前数据集大小较小,后续会逐渐增加
联系作者
email: hanversion@outlook.com
github: GitHub
some-rp-v2This is a filtered subset of lemonilia/Elliquiy-Role-Playing-Forums_2023-04 that's been converted into a typical turn-based conversation format (the added conversations column). These are all 2-user chats, and the first user was assigned to human and the second to gpt.
someone
someone* 订单筛选结果
来源:/home/GRQ/lw_9/ordermsg_logs_20260910。每单两份文件:<日期>_<order_id>.llm_record.jsonl(该单全部模型调用,按时间排序)、<日期>_<order_id>.app.log(该单全部运行日志行)。共 58 单。
订单
日期
产品
回复轮
detect 次
跨度(分)
日志行
20260910_someone1125_test1_2694
20260910
psyche-link
13
8
7.2
392
20260910_someone115_test1_6967
20260910
psyche-link
10
7
6.4
319
20260910_someone1186_test1_1252
20260910
psyche-link
12
7
7.3
356
20260910_someone1286_test1_562
20260910
psyche-link
8
0
6.1
151… See the full description on the dataset page: https://huggingface.co/datasets/HYGGEhygge/someone.sometimesanotion__IF-reasoning-experiment-80-details
Dataset Card for Evaluation run of sometimesanotion/IF-reasoning-experiment-80
Dataset automatically created during the evaluation run of model sometimesanotion/IF-reasoning-experiment-80
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__IF-reasoning-experiment-80-details.somewhereinblog-article
Somewhereinblog Article Archive
Overview
This repository contains a large-scale text dataset scraped from m.somewhereinblog.net, the largest and first-ever Bengali community blogging platform. The primary goal of this archive is to preserve a massive collection of purely human-written blog posts, personal stories, socio-political opinions, and community discussions, creating a distinct record of human-authored text separate from AI-generated content.… See the full description on the dataset page: https://huggingface.co/datasets/sayurio/somewhereinblog-article.Chinese-news-summerysometimesanotion__Lamarck-14B-v0.7-Fusion-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.7-Fusion
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.7-Fusion
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.7-Fusion-details.some-genomessometimesanotion__Lamarck-14B-v0.3-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.3
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.3-details.sometimesanotion__lamarck-14b-prose-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-prose-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-prose-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-prose-model_stock-details.sometimesanotion__lamarck-14b-reason-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-reason-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-reason-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-reason-model_stock-details.DreadPoor__Something-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Something-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Something-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Something-8B-Model_Stock-details.sometimesanotion__Lamarck-14B-v0.4-Qwenvergence-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.4-Qwenvergence
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.4-Qwenvergence
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.4-Qwenvergence-details.sometimesanotion__Qwenvergence-14B-v15-Prose-MS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v15-Prose-MS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v15-Prose-MS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v15-Prose-MS-details.rhea_personsometimesanotion__Qwentinuum-14B-v2-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v2
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v2-details.
