finbert
Datasets
All datasets matching “finbert”finbert-financial-news-sentiment-dataset
📈 Financial News Sentiment Dataset (FinBERT Powered)
Welcome to the official data repository of Lumen Models. This dataset provides a real-time, high-frequency stream of global financial news headlines aggregated from major economic outlets, processed with state-of-the-art Natural Language Processing (NLP).
Every headline is automatically analyzed using FinBERT (a BERT model specifically trained and fine-tuned for financial text analysis) to determine market sentiment with… See the full description on the dataset page: https://huggingface.co/datasets/lumen-models/finbert-financial-news-sentiment-dataset.Sentiments-FinBERT-PT-BR
Dataset
A manually annotated dataset was created to enable supervised training for the FinBERT-PT-BR model, which focuses on sentiment analysis of Brazilian Portuguese financial texts.
More than 1.4 million financial news texts in Portuguese were collected and used for the initial language modeling phase. From this corpus, a sample of 1,000 texts was manually annotated with sentiment labels.
Annotation Process
Three annotators participated in the process.
All texts were… See the full description on the dataset page: https://huggingface.co/datasets/lucas-leme/Sentiments-FinBERT-PT-BR.FinBERT-financial-news-data
FinBERT-financial-news-data
The sentence-level sentiment training data behind
gamug/FinBERT-financial-news --
5,865 sentences from real, English-language financial-news articles (2010s-2020s), each
labeled positive/negative/neutral from an investor/price-impact perspective by an LLM
(DeepSeek deepseek-chat, temperature 0, one sentence at a time, in isolation).
Why this exists
Published as the direct counterpart to the model it trains, so the model's own claims… See the full description on the dataset page: https://huggingface.co/datasets/gamug/FinBERT-financial-news-data.finbert_datasetfinbert_datasetfinbert-sentiment-dataset
