CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01somethingone /MULocBenchThis repository hosts MULocBench, a comprehensive dataset. MULocBench addresses limitations in existing benchmarks by focusing on accurate project localization (e.g., files and functions) for issue resolution, which is a critical first step in software maintenance. It comprises 1,100 issues from 46 popular GitHub Python projects, offering greater diversity in issue types, root causes, location scopes, and file types compared to prior datasets. This dataset provides a more realistic testbed for… See the full description on the dataset page: https://huggingface.co/datasets/somethingone/MULocBench.documenttext-retrieval1K<n<10K0 likes366 downloads6mo agoHugging Face02olarian /something-something-v2document100K<n<1M0 likes248 downloads10mo agoHugging Face03emirgocen /Something-Something-v2text100K<n<1M0 likes86 downloads2y agoHugging Face04someone13574 /smoltalk-binidxtext1M<n<10M1 likes56 downloads2y agoHugging Face05open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v12-Prose-DS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v12-Prose-DS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v12-Prose-DS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details.tabular10K<n<100K0 likes50 downloads2y agoHugging Face06TIME20030221 /something-something-v2document100K<n<1M0 likes49 downloads11d agoHugging Face07open-llm-leaderboard /sometimesanotion__Qwen-2.5-14B-Virmarckeoso-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwen-2.5-14B-Virmarckeoso Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-2.5-14B-Virmarckeoso The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details.tabular10K<n<100K0 likes45 downloads2y agoHugging Face08open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details.tabular10K<n<100K0 likes45 downloads2y agoHugging Face09open-llm-leaderboard /sometimesanotion__Qwentinuum-14B-v7-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v7 Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v7 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v7-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face10open-llm-leaderboard /sometimesanotion__Qwen-14B-ProseStock-v4-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwen-14B-ProseStock-v4 Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-14B-ProseStock-v4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-14B-ProseStock-v4-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face11open-llm-leaderboard /sometimesanotion__LamarckInfusion-14B-v2-detailsgated Dataset Card for Evaluation run of sometimesanotion/LamarckInfusion-14B-v2 Dataset automatically created during the evaluation run of model sometimesanotion/LamarckInfusion-14B-v2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__LamarckInfusion-14B-v2-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face12Siddhu077 /some FinEE Dataset Dataset Description A comprehensive dataset for training financial entity extraction models on Indian banking messages. Contains 152,000+ samples covering SMS, emails, and transaction notifications from major Indian banks. Languages English (en) - 86% Hindi (hi) - 3% Tamil (ta) - 3% Telugu (te) - 3% Bengali (bn) - 3% Kannada (kn) - 2% Supported Transaction Types UPI payments (PhonePe, GPay, Paytm)… See the full description on the dataset page: https://huggingface.co/datasets/Siddhu077/some.texttoken-classification100K<n<1M0 likes42 downloads2mo agoHugging Face13open-llm-leaderboard /sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock Dataset automatically created during the evaluation run of model sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face14open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v13-Prose-DS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v13-Prose-DS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v13-Prose-DS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face15Hanversion /Tieba-SomeInteresting 中文数据集(基于百度贴吧) 简介 本数据集数据来自于2025年3月5日百度贴吧 数据集使用了“孙笑川吧”、“弱智吧”、“中国人口吧”和“航空母舰吧” 数据集问题来自于发帖的标题,答案来自于最热门的回复 思考链来自于 DeepSeek-v3 生成 目前数据集大小较小,后续会逐渐增加 联系作者 email: hanversion@outlook.com github: GitHub text1K<n<10K0 likes32 downloads2y agoHugging Face16ToastyPigeon /some-rp-v2This is a filtered subset of lemonilia/Elliquiy-Role-Playing-Forums_2023-04 that's been converted into a typical turn-based conversation format (the added conversations column). These are all 2-user chats, and the first user was assigned to human and the second to gpt. tabular1K<n<10K2 likes29 downloads1y agoHugging Face17HYGGEhygge /someone someone* 订单筛选结果 来源:/home/GRQ/lw_9/ordermsg_logs_20260910。每单两份文件:<日期>_<order_id>.llm_record.jsonl(该单全部模型调用,按时间排序)、<日期>_<order_id>.app.log(该单全部运行日志行)。共 58 单。 订单 日期 产品 回复轮 detect 次 跨度(分) 日志行 20260910_someone1125_test1_2694 20260910 psyche-link 13 8 7.2 392 20260910_someone115_test1_6967 20260910 psyche-link 10 7 6.4 319 20260910_someone1186_test1_1252 20260910 psyche-link 12 7 7.3 356 20260910_someone1286_test1_562 20260910 psyche-link 8 0 6.1 151… See the full description on the dataset page: https://huggingface.co/datasets/HYGGEhygge/someone.tabular1K<n<10K0 likes24 downloads3d agoHugging Face18open-llm-leaderboard /sometimesanotion__IF-reasoning-experiment-80-detailsgated Dataset Card for Evaluation run of sometimesanotion/IF-reasoning-experiment-80 Dataset automatically created during the evaluation run of model sometimesanotion/IF-reasoning-experiment-80 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__IF-reasoning-experiment-80-details.tabular10K<n<100K1 likes23 downloads2y agoHugging Face19sayurio /somewhereinblog-article Somewhereinblog Article Archive Overview This repository contains a large-scale text dataset scraped from m.somewhereinblog.net, the largest and first-ever Bengali community blogging platform. The primary goal of this archive is to preserve a massive collection of purely human-written blog posts, personal stories, socio-political opinions, and community discussions, creating a distinct record of human-authored text separate from AI-generated content.… See the full description on the dataset page: https://huggingface.co/datasets/sayurio/somewhereinblog-article.imagetext-generation10K<n<100K1 likes23 downloads6mo agoHugging Face20somebreeze /Chinese-news-summerytext100K<n<1M2 likes16 downloads1y agoHugging Face21open-llm-leaderboard /sometimesanotion__Lamarck-14B-v0.7-Fusion-detailsgated Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.7-Fusion Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.7-Fusion The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.7-Fusion-details.tabular10K<n<100K0 likes13 downloads2y agoHugging Face22gonzalobenegas /some-genomestext10M<n<100M0 likes13 downloads1y agoHugging Face23open-llm-leaderboard /sometimesanotion__Lamarck-14B-v0.3-detailsgated Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.3 Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.3-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face24open-llm-leaderboard /sometimesanotion__lamarck-14b-prose-model_stock-detailsgated Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-prose-model_stock Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-prose-model_stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-prose-model_stock-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face25open-llm-leaderboard /sometimesanotion__lamarck-14b-reason-model_stock-detailsgated Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-reason-model_stock Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-reason-model_stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-reason-model_stock-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face26open-llm-leaderboard /DreadPoor__Something-8B-Model_Stock-detailsgated Dataset Card for Evaluation run of DreadPoor/Something-8B-Model_Stock Dataset automatically created during the evaluation run of model DreadPoor/Something-8B-Model_Stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Something-8B-Model_Stock-details.tabular10K<n<100K0 likes11 downloads2y agoHugging Face27open-llm-leaderboard /sometimesanotion__Lamarck-14B-v0.4-Qwenvergence-detailsgated Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.4-Qwenvergence Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.4-Qwenvergence The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.4-Qwenvergence-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face28open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v15-Prose-MS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v15-Prose-MS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v15-Prose-MS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v15-Prose-MS-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face29someoneskilled /rhea_persontextn<1K0 likes9 downloads3y agoHugging Face30open-llm-leaderboard /sometimesanotion__Qwentinuum-14B-v2-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v2 Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v2-details.tabular10K<n<100K0 likes9 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.