datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fomc-events
FOMC Events
Everything the Federal Open Market Committee made public — and, for each of it,
the instant it became public.
571 information arrivals · 184 meetings · 1 862 point-in-time rows ·
1 530 dots · 1 599 votes · 1.38 million words of policy text ·
January 2007 to today, plus the meetings already scheduled to December 2027
The pipeline lives in recipe/ at the same revision as the data.
See PIPELINE.md for the method.
A meeting is not an event
Nothing becomes… See the full description on the dataset page: https://huggingface.co/datasets/ZipLime/fomc-events.fomc_communication
Label Interpretation
LABEL_2: NeutralLABEL_1: HawkishLABEL_0: Dovish
Citation and Contact Information
Cite
Please cite our paper if you use any code, data, or models.
@inproceedings{shah-etal-2023-trillion,
title = "Trillion Dollar Words: A New Financial Dataset, Task {\&} Market Analysis",
author = "Shah, Agam and
Paturi, Suvan and
Chava, Sudheer",
booktitle = "Proceedings of the 61st Annual Meeting of the Association for… See the full description on the dataset page: https://huggingface.co/datasets/gtfintechlab/fomc_communication.fomcfomc-meeting-transcripts
FOMC Meeting Transcripts (1976–2020)
Full-text transcripts of 373 Federal Open Market Committee (FOMC) meetings, from March 1976 through December 2020, converted from the official PDF transcripts published by the Federal Reserve Board.
The FOMC is the body of the U.S. Federal Reserve System that sets monetary policy (the federal funds rate target, balance-sheet policy, etc.). Verbatim meeting transcripts are released to the public with a roughly five-year lag, which is why… See the full description on the dataset page: https://huggingface.co/datasets/brishen/fomc-meeting-transcripts.fomc-personas
FOMC Personas
A speaker-attributed, temporally-resolved corpus of U.S. Federal Open Market Committee (FOMC) members'
public statements, designed for retrieval-augmented digital-twin personas. It accompanies the
paper "A Persona-Based Rate-Action Index" and the code at
github.com/helivan-research/fomc-personas.
The personas power an interactive site: federalreserve.ai.
24,333 chunks across 17 of 19 sitting members (7 Board governors + 10 regional presidents),
spanning 2006–2026… See the full description on the dataset page: https://huggingface.co/datasets/helivan/fomc-personas.fomc-communicationDataset adapted from original work by Shah et al.
About Dataset
The dataset is a collection of sentences from FOMC speeches, meeting minutes and press releases (see corresponding paper). A subset of the data has been manually annotated as hawkish, dovish, or neutral.
Label mapping
LABEL 2: Neutral
LABEL 1: Hawkish
LABEL 0: Dovish
fomc-communication-counterfactualDataset adapted from original work by Shah et al.
About Dataset
The dataset is a collection of sentences from FOMC speeches, meeting minutes and press releases (see corresponding paper). A subset of the data has been manually annotated as hawkish, dovish, or neutral.
Label mapping
LABEL 2: Neutral
LABEL 1: Hawkish
LABEL 0: Dovish
Counterfactual generation split
Additionally, for counterfactual generation tasks, we add a custom split with target classes in… See the full description on the dataset page: https://huggingface.co/datasets/TextCEsInFinance/fomc-communication-counterfactual.fomc-draft-v0
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/suschi1993/fomc-draft-v0.fomcFOMC-counterfactuals
