datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
financial-excel-modeling-sftreward-modeling-short-tokenized
Dataset Card for "reward-modeling-short-tokenized"
More Information needed
c4_newslike_url_onlydesign-bench
SciModelingBench Design-Bench Data
Canonical, provenance-tracked observations for scientific modeling and design Tasks.
GitHub
·
Python Package
·
Documentation
·
Organization
This repository stores the scientific observation layer used by the
SciModelingBench Design-Bench suite. The Python package supplies validators,
Agent-visible Protocols, trusted Objectives, submission contracts, and Task
metrics. Data and evaluation logic… See the full description on the dataset page: https://huggingface.co/datasets/sci-modeling-bench/design-bench.financial-statement-modeling-sft-dpo-2026
📈 Enterprise Financial AI, SEC 10-K & Valuation Modeling SFT/DPO Dataset (2026)
High-precision multi-turn instruction tuning and preference optimization dataset with step-by-step arithmetic Chain-of-Thought (<thought>) reasoning chains for fine-tuning LLMs (Llama-3.3, Qwen-2.5-Coder, DeepSeek-R1-Distill, Mistral) into Wall Street Equity Research Associates, M&A Valuation Modelers, and Senior Forensic Auditors.
📊 Dataset Architecture & Highlights
Multi-Turn… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/financial-statement-modeling-sft-dpo-2026.modeling_datared_teaming_reward_modeling_pairwise
Dataset Card for "red_teaming_reward_modeling_pairwise"
More Information needed
reward-modeling-long-tokenized
Dataset Card for "reward-modeling-long-tokenized"
More Information needed
red_teaming_reward_modeling_pairwise_no_as_an_ai
Dataset Card for "red_teaming_reward_modeling_pairwise_no_as_an_ai"
More Information needed
sharegpt_reward_modeling_pairwise_no_as_an_ai
Dataset Card for "sharegpt_reward_modeling_pairwise_no_as_an_ai"
More Information needed
reward_modeling_dataset
Dataset Card for "reward_modeling_dataset"
More Information needed
gpteacher_reward_modeling_pairwise
Dataset Card for "gpteacher_reward_modeling_pairwise"
More Information needed
modeling
Dataset Card for "modeling"
More Information needed
sharegpt_reward_modeling_pairwise
Dataset Card for "sharegpt_reward_modeling_pairwise"
More Information needed
ilm_deuplift-modeling-synthetic-benchmark
Synthetic uplift benchmark with known ground-truth treatment effect
100,000 training rows, 20,000 validation rows, generated for
uplift-modeling's Gate 0: checking that
meta-learners (S/T/X-learner) actually recover a real treatment effect before trusting them
on data where no individual ground truth is ever available - which is true of essentially
all real causal-inference data, by the fundamental problem of causal inference (nobody
observes both potential outcomes for the same… See the full description on the dataset page: https://huggingface.co/datasets/Bauxitiego/uplift-modeling-synthetic-benchmark.Churnn_modeling_datasetilm_esilm_itachurn_modeling_datasetChurn_modeling_datasetchurn_modelingarxiv_topic_modelingRAG-Reward-Modeling-v2
Dataset Card for "RAG-Reward-Modeling-v2"
More Information needed
ilm_poldolly_reward_modeling_pairwise
Dataset Card for "dolly_reward_modeling_pairwise"
More Information needed
modeling_v1
Dataset Card for "modeling_v1"
More Information needed
ilm_euamazon_reviews_user_modelingru_language_modeling_v4
