datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
financial-economics-reasoning
Model Card
📌 Summary
financial-economics-reasoning dataset was constructed using advanced Inference Distillation techniques. We employed the qwen-3-235b-a22b-thinking-2507 model as the Teacher Model to process the open-source BAAI/IndustryInstruction_Finance-Economics dataset, which contains 122,378 bilingual (Chinese-English) entries in finance, economics, and business.
Unlike standard distillation datasets that only provide final answers, this dataset retains the… See the full description on the dataset page: https://huggingface.co/datasets/Jackrong/financial-economics-reasoning.economist-tui-sessions
Coding agent session traces for thomasmustier/economist-tui-sessions
This dataset contains redacted coding agent session traces collected while working on tmustier/economist-tui. The traces were exported with pi-share-hf from a local pi workspace and filtered to keep only sessions that passed deterministic redaction and LLM review.
Data description
Each *.jsonl file is a redacted pi session. Sessions are stored as JSON Lines files where each line is a structured… See the full description on the dataset page: https://huggingface.co/datasets/thomasmustier/economist-tui-sessions.EconSafeBench
Dataset Card for EconSafeBench
EconSafeBench evaluates the safety of LLM agents in executable economic
environments, testing whether agents violate regulatory, informational,
fairness, or data-use constraints while pursuing an economic objective
under three distinct sources of pressure.
Dataset Details
Dataset Description
EconSafeBench contains 828 cases spanning five executable economic
scenarios and four categories of safety violations. Unlike… See the full description on the dataset page: https://huggingface.co/datasets/Yuzhu0921/EconSafeBench.stata-econ-bench
Stata Econometrics Benchmark
250 natural-language econometrics & statistics tasks, each solvable with a short Stata program and graded by executing the generated code against 1250 hidden numeric test cases (5 per problem). This is an execution-based benchmark: a solution is correct only if running it reproduces the expected numeric result within a per-case tolerance — not by string match.
At a glance
250 problems, 1250 test cases (5 per problem)
Target language:… See the full description on the dataset page: https://huggingface.co/datasets/eltokh7/stata-econ-bench.instruct-economics-pashto
Instruct Economics Pashto
This dataset contains Pashto translations of high‑quality economics instruction datasets.It is designed for fine‑tuning conversational Large Language Models (LLMs) on advanced economic reasoning, welfare theory, micro/macro analysis, and policy evaluation — all in the Pashto language.
دا ډیټاسیټ د اقتصاد د لوړو مفاهیمو، هوساینې تیورۍ، مایکرو او ماکرو اقتصاد، او د ټولنیزو پالیسیو د تحلیل لپاره د لارښوونې ډیالوګونو پښتو ژباړې لري. دا د Pashto ژبې لپاره د… See the full description on the dataset page: https://huggingface.co/datasets/nassimjp/instruct-economics-pashto.asotele-eval-nigerian-economy
Asotele Eval — Nigerian Economic Reasoning (v1)
A small, hand-curated rubric-graded evaluation set for measuring whether a language model can reason about the Nigerian economy the way an experienced Nigerian credit officer, SME owner, or independent analyst would.
This is v1 (seed), intentionally small. Each record is dense, with citations the model must use, omissions that lose points, a reference answer, and a per-record scoring rubric. The goal is to surface qualitative reasoning… See the full description on the dataset page: https://huggingface.co/datasets/Apexgridapps/asotele-eval-nigerian-economy.
