CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01piergiuliol /financial-excel-modeling-sfttext1K<n<10K2 likes3.5k downloads5mo agoHugging Face02andersonbcdefg /reward-modeling-short-tokenized Dataset Card for "reward-modeling-short-tokenized" More Information needed 100K<n<1M2 likes341 downloads3y agoHugging Face03bs-modeling-metadata /c4_newslike_url_onlytext10M<n<100M0 likes311 downloads5y agoHugging Face04sci-modeling-bench /design-bench SciModelingBench Design-Bench Data Canonical, provenance-tracked observations for scientific modeling and design Tasks. GitHub &nbsp;·&nbsp; Python Package &nbsp;·&nbsp; Documentation &nbsp;·&nbsp; Organization This repository stores the scientific observation layer used by the SciModelingBench Design-Bench suite. The Python package supplies validators, Agent-visible Protocols, trusted Objectives, submission contracts, and Task metrics. Data and evaluation logic… See the full description on the dataset page: https://huggingface.co/datasets/sci-modeling-bench/design-bench.tabular1M<n<10M0 likes267 downloads2mo agoHugging Face05beatsprom /financial-statement-modeling-sft-dpo-2026 📈 Enterprise Financial AI, SEC 10-K & Valuation Modeling SFT/DPO Dataset (2026) High-precision multi-turn instruction tuning and preference optimization dataset with step-by-step arithmetic Chain-of-Thought (<thought>) reasoning chains for fine-tuning LLMs (Llama-3.3, Qwen-2.5-Coder, DeepSeek-R1-Distill, Mistral) into Wall Street Equity Research Associates, M&A Valuation Modelers, and Senior Forensic Auditors. 📊 Dataset Architecture & Highlights Multi-Turn… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/financial-statement-modeling-sft-dpo-2026.texttext-generationn<1K0 likes253 downloads28d agoHugging Face06ADS599-Capstone /modeling_datatabular10M<n<100M0 likes218 downloads6mo agoHugging Face07andersonbcdefg /red_teaming_reward_modeling_pairwise Dataset Card for "red_teaming_reward_modeling_pairwise" More Information needed text10K<n<100K7 likes189 downloads3y agoHugging Face08andersonbcdefg /reward-modeling-long-tokenized Dataset Card for "reward-modeling-long-tokenized" More Information needed 100K<n<1M1 likes147 downloads3y agoHugging Face09andersonbcdefg /red_teaming_reward_modeling_pairwise_no_as_an_ai Dataset Card for "red_teaming_reward_modeling_pairwise_no_as_an_ai" More Information needed text10K<n<100K6 likes135 downloads3y agoHugging Face10andersonbcdefg /sharegpt_reward_modeling_pairwise_no_as_an_ai Dataset Card for "sharegpt_reward_modeling_pairwise_no_as_an_ai" More Information needed text10K<n<100K3 likes102 downloads3y agoHugging Face11HumanDynamics /reward_modeling_dataset Dataset Card for "reward_modeling_dataset" More Information needed text10K<n<100K2 likes81 downloads3y agoHugging Face12andersonbcdefg /gpteacher_reward_modeling_pairwise Dataset Card for "gpteacher_reward_modeling_pairwise" More Information needed text1K<n<10K2 likes71 downloads3y agoHugging Face13Gae8J /modeling Dataset Card for "modeling" More Information needed audioaudio-classificationn<1K0 likes68 downloads3y agoHugging Face14andersonbcdefg /sharegpt_reward_modeling_pairwise Dataset Card for "sharegpt_reward_modeling_pairwise" More Information needed text10K<n<100K1 likes53 downloads3y agoHugging Face15interlinguistic-language-modeling /ilm_detext1M<n<10M0 likes53 downloads6mo agoHugging Face16Bauxitiego /uplift-modeling-synthetic-benchmark Synthetic uplift benchmark with known ground-truth treatment effect 100,000 training rows, 20,000 validation rows, generated for uplift-modeling's Gate 0: checking that meta-learners (S/T/X-learner) actually recover a real treatment effect before trusting them on data where no individual ground truth is ever available - which is true of essentially all real causal-inference data, by the fundamental problem of causal inference (nobody observes both potential outcomes for the same… See the full description on the dataset page: https://huggingface.co/datasets/Bauxitiego/uplift-modeling-synthetic-benchmark.tabulartabular-regression100K<n<1M0 likes50 downloads1mo agoHugging Face17vishesh19 /Churnn_modeling_datasettabular10K<n<100K0 likes41 downloads6d agoHugging Face18interlinguistic-language-modeling /ilm_estext1M<n<10M0 likes40 downloads6mo agoHugging Face19interlinguistic-language-modeling /ilm_itatext1M<n<10M0 likes39 downloads4mo agoHugging Face20Warrior1110 /churn_modeling_datasettabular10K<n<100K0 likes38 downloads6d agoHugging Face21vishesh19 /Churn_modeling_datasettext10K<n<100K0 likes35 downloads6d agoHugging Face22indxr0101 /churn_modelingtabular10K<n<100K0 likes35 downloads6d agoHugging Face23bunkalab /arxiv_topic_modelingtext1K<n<10K4 likes33 downloads2y agoHugging Face24HanningZhang /RAG-Reward-Modeling-v2 Dataset Card for "RAG-Reward-Modeling-v2" More Information needed text10K<n<100K0 likes31 downloads1y agoHugging Face25interlinguistic-language-modeling /ilm_poltext1M<n<10M0 likes31 downloads4mo agoHugging Face26andersonbcdefg /dolly_reward_modeling_pairwise Dataset Card for "dolly_reward_modeling_pairwise" More Information needed text10K<n<100K2 likes26 downloads3y agoHugging Face27Gae8J /modeling_v1 Dataset Card for "modeling_v1" More Information needed audion<1K0 likes22 downloads3y agoHugging Face28interlinguistic-language-modeling /ilm_eutext1M<n<10M0 likes21 downloads6mo agoHugging Face29maveriq /amazon_reviews_user_modelingtext1K<n<10K0 likes20 downloads2y agoHugging Face30patrikgerard /ru_language_modeling_v4text1K<n<10K0 likes20 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.