CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Mumon /mmlu-pro-self-cot-deepseek-r1Use deepseek-r1 to generate COT in few-shot examples. tabular10K<n<100K1 likes1.7k downloads2y agoHugging Face02Rapidata /text-2-video-human-preferences-seedance-1-pro Rapidata Video Generation Seedance 1 Pro Human Preference In this dataset, ~60k human responses from ~20k human annotators were collected to evaluate Seedance 1 Pro video generation model on our benchmark. This dataset was collected in roughtly 30 min using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation. Explore our latest model rankings on our website. If you get value from this dataset and would like to see more in the future, please… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/text-2-video-human-preferences-seedance-1-pro.imagevideo-classification1K<n<10K9 likes708 downloads1y agoHugging Face03letrinhan /vn-provinces-criminal-cases-prosecuted Vietnam criminal cases prosecuted Vietnam criminal cases prosecuted. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system). Figures Hero Comparison Color key Files provinces (189 rows) data/provinces.csv data/provinces.dta data/provinces.xlsx regions (18 rows) data/regions.csv data/regions.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-prosecuted.tabularn<1K0 likes98 downloads5d agoHugging Face04PocketDoc /Dans-Prosemaxx-Adventuretabularn<1K4 likes74 downloads2y agoHugging Face05timchen0618 /browsecomp-plus-sel-tools-test300-gemini-3p1-pro-v1tabularn<1K0 likes64 downloads4mo agoHugging Face06open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v12-Prose-DS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v12-Prose-DS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v12-Prose-DS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details.tabular10K<n<100K0 likes50 downloads2y agoHugging Face07hmcgovern /de-en-prosetabular1M<n<10M0 likes47 downloads2y agoHugging Face08open-llm-leaderboard /sometimesanotion__Qwen-14B-ProseStock-v4-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwen-14B-ProseStock-v4 Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-14B-ProseStock-v4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-14B-ProseStock-v4-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face09open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v13-Prose-DS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v13-Prose-DS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v13-Prose-DS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face10wolfvswhale /prose-cadence-stats Prose cadence statistics Measurements of 38 stylometric features across 5,402 documents, split by authorship (human or machine) and by register (informal, formal, multi-paragraph). There is no text in this dataset. Every row is a set of numbers plus a stable reference to the document it was measured from. That is deliberate, and both reasons matter. The sources carry incompatible licenses, so republishing a merged text corpus would be a mess. Measurements are facts about text… See the full description on the dataset page: https://huggingface.co/datasets/wolfvswhale/prose-cadence-stats.tabulartext-classification1K<n<10K0 likes40 downloads2mo agoHugging Face11woog /arena-prose-100-49-models Arena Prose: 100 prompts × 50 models A paired exploratory AI-text-detection corpus: 5,000 successful generated responses from 50 models, each answering the same 100 English prose prompts. Generation was performed through OpenRouter in September 2026 with optional reasoning disabled and mandatory reasoning set to low. This is an independent local benchmark inspired by Pangram 4 §5.2, not an official Pangram dataset or exact replication. Loading from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/woog/arena-prose-100-49-models.tabulartext-generation1K<n<10K0 likes40 downloads2d agoHugging Face12zarahall /prose-steering-results-n32tabularn<1K0 likes38 downloads1y agoHugging Face13artindnr /khayyam-challenge-prose-terra Khayyam Challenge - Prose (Terra) AI-generated Persian prose descriptions of the 20 classical poems in the Khayyam Challenge benchmark, produced by the model internally labeled Terra. Split low / medium / long by poem length (low: 10, medium: 7, long: 3). Each record contains everything in the poems repo (id, poet, title, form, verse_count, theme, text, ...) plus a conversion object with the generated prose, so this repo is self-contained -- no join required: from datasets… See the full description on the dataset page: https://huggingface.co/datasets/artindnr/khayyam-challenge-prose-terra.tabularn<1K1 likes34 downloads2mo agoHugging Face14artindnr /khayyam-challenge-prose-luna Khayyam Challenge - Prose (Luna) AI-generated Persian prose descriptions of the 20 classical poems in the Khayyam Challenge benchmark, produced by the model internally labeled Luna. Split low / medium / long by poem length (low: 10, medium: 7, long: 3). Each record contains everything in the poems repo (id, poet, title, form, verse_count, theme, text, ...) plus a conversion object with the generated prose, so this repo is self-contained -- no join required: from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/artindnr/khayyam-challenge-prose-luna.tabularn<1K1 likes32 downloads2mo agoHugging Face15khursanirevo /sft-bm-prose khursanirevo/sft-bm-prose Bahasa Melayu prose-format text (long-form lessons + textbook-style, ~232k rows). Splits split rows train 221,170 validation 11,635 Stratified 95/5 by source/category (seed=42). Source files data/midtrain/synth_hf_prose.jsonl data/midtrain/synth_bm_50m.jsonl Schema Each row is a JSON object. See the loader script for field details. Provenance Generated as part of MaLLaM 2026… See the full description on the dataset page: https://huggingface.co/datasets/khursanirevo/sft-bm-prose.tabulartext-generation100K<n<1M0 likes18 downloads3mo agoHugging Face16joyfine /router_SFT_self_generated_data_mmlu_pro_science_OLMo-2-1124-13B-Instructtabular1K<n<10K0 likes14 downloads5mo agoHugging Face17timchen0618 /browsecomp-plus-sel-tools-test300-gemini-2p5-pro-v1tabularn<1K0 likes14 downloads4mo agoHugging Face18open-llm-leaderboard /sometimesanotion__lamarck-14b-prose-model_stock-detailsgated Dataset Card for Evaluation run of sometimesanotion/lamarck-14b-prose-model_stock Dataset automatically created during the evaluation run of model sometimesanotion/lamarck-14b-prose-model_stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__lamarck-14b-prose-model_stock-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face19semran1 /megamath-web-protabular10M<n<100M1 likes12 downloads1y agoHugging Face20joyfine /router_SFT_self_generated_data_mmlu_pro_science_Meta-Llama-3-8B-Instructtabular1K<n<10K0 likes11 downloads5mo agoHugging Face21joyfine /router_SFT_self_generated_data_mmlu_pro_science_Qwen3-4B_aimetabularn<1K0 likes11 downloads5mo agoHugging Face22open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v15-Prose-MS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v15-Prose-MS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v15-Prose-MS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v15-Prose-MS-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face23zarahall /prose-steering-results-n35tabularn<1K0 likes10 downloads1y agoHugging Face24joyfine /router_SFT_self_generated_data_mmlu_pro_science_Qwen3-0.6Btabular1K<n<10K0 likes9 downloads6mo agoHugging Face25joyfine /router_SFT_self_generated_data_mmlu_pro_science_Qwen3-8Btabular1K<n<10K0 likes9 downloads6mo agoHugging Face26open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v3-Prose-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v3-Prose Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v3-Prose The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v3-Prose-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face27open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v6-Prose-model_stock-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v6-Prose-model_stock Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v6-Prose-model_stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v6-Prose-model_stock-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face28open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v6-Prose-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v6-Prose Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v6-Prose The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v6-Prose-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face29zarahall /prose-steering-results-n34tabularn<1K0 likes8 downloads1y agoHugging Face30zarahall /prose-steering-results-n325tabularn<1K0 likes8 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.