finra
Datasets
All datasets matching “finra”FinRAGBench-V
FinRAGBench-V: A Benchmark for Multimodal RAG with Visual Citation in the Financial Domain 🤗 Code 📄 Paper
Overview
FinRAGBench-V is a comprehensive benchmark for visual retrieval-augmented generation (RAG) in finance, addressing the challenge that most existing financial RAG research focuses predominantly on text while overlooking rich visual content in financial documents. By integrating multimodal data and providing visual citation, FinRAGBench-V ensures traceability… See the full description on the dataset page: https://huggingface.co/datasets/zhaosuifeng/FinRAGBench-V.Fin-RATE
📝 Fin-RATE: Financial Analytics and Tracking Evaluation Benchmark for LLMs on SEC Filings
Fin-RATE is a real-world benchmark to evaluate large language models (LLMs) on professional-grade reasoning over U.S. SEC filings.
It targets financial analyst workflows that demand:
📄 Long-context understanding
⏱️ Cross-year tracking
🏢 Cross-company comparison
📊 Structured diagnosis of model failures
📘 [Paper (arXiv link TBD)] | 🤗 Dataset
⬇️ SEC-based QA benchmark with 7,500… See the full description on the dataset page: https://huggingface.co/datasets/GGLabYale/Fin-RATE.finra-brokercheck-scraper
FINRA BrokerCheck Scraper · Advisors, Firms & Disclosures
Scrape financial advisors, firm affiliations, CRDs, registration scope, and disclosure histories directly from FINRA BrokerCheck API into clean dataset rows.
Rows in this dataset
1,430
Fields
22
Collector runs behind it
50
Most recent observation
2026-08-03
What this is
Every row here was returned by a real run of a public collector. Nothing is generated from a
template over a… See the full description on the dataset page: https://huggingface.co/datasets/reapxdev/finra-brokercheck-scraper.FinRAG
FinRAG
Dataset Description
This dataset contains 12,500 financial reasoning questions based on real-world financial documents, earnings reports, and financial tables. Each question is accompanied by a correct answer and four carefully crafted distractor answers, making it suitable for multiple-choice question answering tasks and assessing financial numerical reasoning capabilities.
Dataset Summary
Total Examples: 12,500
Format: Multiple-choice questions with 5… See the full description on the dataset page: https://huggingface.co/datasets/trismik/FinRAG.FinRAG-GRPO
FinRAG-GRPO Preference Dataset
A Chinese-language preference dataset for training Reasoning Reward Models (ReasRM) via GRPO-based reinforcement learning.
🚧 This dataset is actively maintained and will be expanded with additional domains and languages over time.
Dataset Summary
This dataset contains pairwise preference samples designed to train a reward model that reasons before judging — the model generates an evaluation rationale before outputting a preference label… See the full description on the dataset page: https://huggingface.co/datasets/SamWang0405/FinRAG-GRPO.Fin-RATE
📝 Fin-RATE: Financial Analytics and Tracking Evaluation Benchmark for LLMs on SEC Filings
Fin-RATE is a real-world benchmark to evaluate large language models (LLMs) on professional-grade reasoning over U.S. SEC filings.
It targets financial analyst workflows that demand:
📄 Long-context understanding
⏱️ Cross-year tracking
🏢 Cross-company comparison
📊 Structured diagnosis of model failures
📘 [Paper (arXiv link TBD)] | 🤗 Dataset
⬇️ SEC-based QA benchmark with 7,500… See the full description on the dataset page: https://huggingface.co/datasets/idleengine/Fin-RATE.
