datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
gspc-ai-economy-index
GSPC — ai adoption components facts (Eurostat)
SWIFT census (live): https://councilof.ai/api/swift
XRPL reader (live): https://councilof.ai/api/xrpl
Live axis name: ai-adoption-components — MEASURED as two Eurostat series (deterministic-facts, n=2). Not an index. No composite formula. No MEASURED-INDEX-v0.1 sticker (C-2026-0826-05: do not restore).
This Hub repo id keeps the legacy slug gspc-ai-economy-index for inbound links only. Cite the live axis name. Do not stamp an… See the full description on the dataset page: https://huggingface.co/datasets/csoai/gspc-ai-economy-index.IndustryInstruction_Finance-Economics
IndustryInstruction: Finance & Economics
This repository contains the IndustryInstruction: Finance & Economics domain subset of BAAI/IndustryInstruction.
Refer to the parent dataset card for data construction, intended use, limitations,
and licensing details.
Citation
If you use this dataset in your work, please cite IndustryInstruction:
@misc{shi2024industryinstruction,
title = {IndustryInstruction},
author = {Xiaofeng Shi and Lulu Zhao and Hua Zhou… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryInstruction_Finance-Economics.financial-economics-reasoning
Model Card
📌 Summary
financial-economics-reasoning dataset was constructed using advanced Inference Distillation techniques. We employed the qwen-3-235b-a22b-thinking-2507 model as the Teacher Model to process the open-source BAAI/IndustryInstruction_Finance-Economics dataset, which contains 122,378 bilingual (Chinese-English) entries in finance, economics, and business.
Unlike standard distillation datasets that only provide final answers, this dataset retains the… See the full description on the dataset page: https://huggingface.co/datasets/Jackrong/financial-economics-reasoning.economist-tui-sessions
Coding agent session traces for thomasmustier/economist-tui-sessions
This dataset contains redacted coding agent session traces collected while working on tmustier/economist-tui. The traces were exported with pi-share-hf from a local pi workspace and filtered to keep only sessions that passed deterministic redaction and LLM review.
Data description
Each *.jsonl file is a redacted pi session. Sessions are stored as JSON Lines files where each line is a structured… See the full description on the dataset page: https://huggingface.co/datasets/thomasmustier/economist-tui-sessions.trade-economy-index
The Trade Economy Index
Release: 2026.1.6
How the US skilled trades fare in the AI era: labor, market structure, unit
economics, cash cycle, AI exposure, geography, and valuation for 13 commercial
trades, with per-cell citations where applicable and table-level provenance for
derived and aggregate tables.
Companion site: tradesindex.org. Published by
Level. Archived with a DOI:
10.5281/zenodo.21762674.
Also available as CSV and JSON on
GitHub.
Why this exists
Most… See the full description on the dataset page: https://huggingface.co/datasets/LevelCFO/trade-economy-index.semantic-economyall things are now lawful to you in jack feist
EA-RHIZOME-SE-01 — semantic economy
The archive's political economy of meaning, seeded at the stolon the collapse body left open. Its roles are NOT the collapse roles: where that body asks what narrows, this one asks who gains by the narrowing.
These are symbola. They are for traversal.
A token broken in two, each half held by a different party, no half carrying complete authority, the fit of the fracture proving the… See the full description on the dataset page: https://huggingface.co/datasets/leesharks/semantic-economy.labour-economy-unmeasured
Labour economy — UNMEASURED, on purpose
Register: UNMEASURED. Absence is not zero.
Three contextual indices — AI-economy · human-labour · humanoid-labour — declared empty until INDEX-METHOD freezes a bank and usable n. They are a contextual firewall and must never be fused into GSPC (SHA-256 / Ed25519) grading cells.
Method: https://github.com/CSOAI-ORG/councilof-ai/blob/master/docs/SOVOS/INDEX-METHOD-0.1.md (branch until merge)
Live API (after master merge): GET… See the full description on the dataset page: https://huggingface.co/datasets/csoai/labour-economy-unmeasured.economy-watchers-survey
economy-watchers-survey
Economy Watchers Survey data.It is automatically updated by GitHub Actions as the economy watcher is updated.The dataset for tasks is retarfi/economy-watchers-survey-evaluation.
景気ウォッチャー調査のデータを自動更新・整形・抽出を行います。自動更新はGitHub Actionsによって月次で行われます。タスク用のデータセットはretarfi/economy-watchers-survey-evaluationから利用可能です。
Data detail
Please refer to the following papers for the data detail.データの詳細は、以下の論文を参照してください。
English paper:… See the full description on the dataset page: https://huggingface.co/datasets/retarfi/economy-watchers-survey.EconSafeBench
Dataset Card for EconSafeBench
EconSafeBench evaluates the safety of LLM agents in executable economic
environments, testing whether agents violate regulatory, informational,
fairness, or data-use constraints while pursuing an economic objective
under three distinct sources of pressure.
Dataset Details
Dataset Description
EconSafeBench contains 828 cases spanning five executable economic
scenarios and four categories of safety violations. Unlike… See the full description on the dataset page: https://huggingface.co/datasets/Yuzhu0921/EconSafeBench.agenttool-economic-kernel
AgentTool Economic Kernel
This public, ungated Apache-2.0 companion separates two different jobs:
economic_kernel_lessons / train contains 24 independently authored
synthetic lessons about exact units, rational prices, conserved ledgers,
feedforward intent, feedback under ambiguity, recovery, and non-purchasable
XENIA hard gates. The publisher admits only these rows for training.
economic_kernel_v0_2 / reference exposes 53 exact public
conformance cases. They are held out from… See the full description on the dataset page: https://huggingface.co/datasets/Yu-and-Ai/agenttool-economic-kernel.igcse-economics-qa-2kEconSkills
EconSkills
EconSkills is a library of 50 reusable, instance-free skills distilled from
verified successful trajectories on EconWebArena.
Each skill is a parameterized standard operating procedure (SOP) for retrieving a
specific kind of live economic figure from an authoritative web portal (central
banks, statistical agencies, market data sites, and government fee or benefit
schedules). Instance-specific values from the source task (for example a date,
country, currency, or… See the full description on the dataset page: https://huggingface.co/datasets/EconWebArena/EconSkills.dataforge-economics
Dataset Card for dataforge-economics
Overview
This dataset, teknium/dataforge-economics, is a specialized collection of 1,000 synthetic examples in the field of economics. It has been generated using OpenAI's GPT-4 and a custom data synthesis pipeline named DataForge, developed by me.
Dataset Description
Data Collection and Synthesis
The data in teknium/dataforge-economics has been synthetically generated using OpenAI's GPT-4 language model. The… See the full description on the dataset page: https://huggingface.co/datasets/teknium/dataforge-economics.EcoNexus-Knowledge
数据集简介
EcoNexus为江苏龙衡环境打造的环保领域专用AI系统,包括EcoNexus-Knowledge环保专用数据集,及EcoNexus-AI环保专业AI大模型系统。
EcoNexus-Knowledge基础版数据量约为70k。
数据集覆盖范围
环境领域相关法律法规、标准、技术规范以及导则等文件
生态环境部典型行政处罚案例
江苏省生态环境厅典型行政处罚案例、咨询回复
后续会持续更新最新内容,包括收录各领域独家经验文档。
igcse-economics-qakv-reuse-econ-traces
KV Reuse Econ Traces — a headline and the closed form that predicts it, side by side
65 per-workload first-touch prefill accounting rows: 36 from a synthetic size ramp, 29 from the
real Mooncake FAST'25 trace.
Why this dataset exists
We published a 90.0% mean first-touch prefill cut. It is an exact arithmetic identity, not a
measured efficiency — and rather than say so in a footnote, this dataset ships the closed form as
a column beside the measurement:… See the full description on the dataset page: https://huggingface.co/datasets/nickh007/kv-reuse-econ-traces.economics
Dataset Card for cxllin/economics
This dataset aims to represent knowledge within the realm of economics
Dataset Details
Featuring Macro, Micro, and Math texbooks
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/cxllin/economics.Nepal_Economic_Statistics_FAQ
🇳🇵 Nepal Economic Statistics FAQ (Nepali) — Merged Dataset README
A merged, sequentially re-indexed dataset of 645 Nepali-language factual Q&A pairs covering three related economic domains: commodity/SITC trade, GDP by ISIC sector, and customs import/export by checkpoint. Built by combining three source datasets into a single JSONL file.
🔖 TL;DR
What
Value
Total records
645
Output file
merged_nepal_economic_stats.jsonl
File size
~1.21 MB… See the full description on the dataset page: https://huggingface.co/datasets/sabin1234/Nepal_Economic_Statistics_FAQ.adaption-econ-finance-qa-pairs
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-econ_finance_qa_pairs
This dataset contains prompt-completion pairs focused on economics, finance, and business calculations. The content covers topics such as growth theories, market metrics, profit margins, and government financial roles. Responses include step-by-step reasoning, numerical computations, and theoretical explanations tailored to specific queries.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/mishface123/adaption-econ-finance-qa-pairs.english-basic-economics-qa-50stata-econ-bench
Stata Econometrics Benchmark
250 natural-language econometrics & statistics tasks, each solvable with a short Stata program and graded by executing the generated code against 1250 hidden numeric test cases (5 per problem). This is an execution-based benchmark: a solution is correct only if running it reproduces the expected numeric result within a per-case tolerance — not by string match.
At a glance
250 problems, 1250 test cases (5 per problem)
Target language:… See the full description on the dataset page: https://huggingface.co/datasets/eltokh7/stata-econ-bench.economy-and-finance
Ekşi Sözlük Türkçe Teknoloji Dataset
Ekşi Sözlük'ten derlenen, teknoloji kategorisine ait Türkçe kullanıcı entry'lerinden oluşan bir veri setidir. Türkçe NLP araştırmaları ve LLM eğitimi için hazırlanmıştır.
İçerik
Yapay zeka, ChatGPT, sosyal medya algoritmaları, kripto para ve diğer teknoloji konularını kapsayan 24 farklı başlık altında toplanmış entry'lerden oluşmaktadır.
Alan
Kapsanan Başlıklar
Yapay Zeka
yapay zeka, chatgpt, claude ai, meta ai, apple… See the full description on the dataset page: https://huggingface.co/datasets/turkish-nlp-datasets/economy-and-finance.basic_economics_questions_ts_test_1Synthethic Question & Answer dataset trained on a corpus of the book Basic Economics by Thomas Sowell.
Formating could be improved, as model trained on this dataset write \n tokens as words and not as newline, so I guess it gets tokenized in a way different from expectations.
Note that prompt format isn't very consistent in every sample.
Spicyboros 7B gguf was used as a model that generated synthetic responses, so it was all generated locally without leaving the device, as opposed to how… See the full description on the dataset page: https://huggingface.co/datasets/adamo1139/basic_economics_questions_ts_test_1.basic_economics_questions_ts_test_2Economics_25k
WithinUsAI/Economics_25k — Master Scholars Academics (25k)
8,000 verification (TRUE/FALSE + correction)
9,000 self-contained quantitative/structured reasoning
8,000 definitions / micro-refreshers
Generated: 2026-01-04T05:12:40Z
basic_economics_questions_ts_test_4basic_economics_questions_ts_test_3ap_economicsinstruct-economics-pashto
Instruct Economics Pashto
This dataset contains Pashto translations of high‑quality economics instruction datasets.It is designed for fine‑tuning conversational Large Language Models (LLMs) on advanced economic reasoning, welfare theory, micro/macro analysis, and policy evaluation — all in the Pashto language.
دا ډیټاسیټ د اقتصاد د لوړو مفاهیمو، هوساینې تیورۍ، مایکرو او ماکرو اقتصاد، او د ټولنیزو پالیسیو د تحلیل لپاره د لارښوونې ډیالوګونو پښتو ژباړې لري. دا د Pashto ژبې لپاره د… See the full description on the dataset page: https://huggingface.co/datasets/nassimjp/instruct-economics-pashto.english-introduction-to-economics-30
