datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bis_central_bank_speeches
Dataset Card
This dataset consists of central bankers speeches from 1997 to 2025 scrapped automatically from the Bank Of International Settlements website.
Each speech is associated to a central bank, a date and a description (a metadata provided by the website).
The dataset covers a wide range of topics in economics from monetary policies to world outlooks, financial stability, unemployment, fiscal policies...
Credits
Full credits to the Bank of International… See the full description on the dataset page: https://huggingface.co/datasets/samchain/bis_central_bank_speeches.sejm-speeches-corpus
Sejm Speeches Corpus — kadencje VII–X
Wersjonowany korpus wypowiedzi z oficjalnych materiałów Sejmu RP,
przygotowany do treningu językowego, walidacji i reprodukowalnych badań.
Object i Version
Object: slayer://object/dataset/piotrsty-sejm-speeches-corpus.
Wydanie: 1.0.1.
Niezmienny alias: v1.0.1 w repozytorium Hugging Face.
Poprzednie niezmienne wydanie: v1.0.0; relacja SUPERSEDES.
Profil: slayer.ai/dataset-release/v1.
Pełny digest wersji oraz payload URI… See the full description on the dataset page: https://huggingface.co/datasets/PiotrSty/sejm-speeches-corpus.german-parliament-speeches
German Parliament Speeches
This dataset contains speeches from the German parliament, derived from the Open Discourse Project (Harvard Dataverse).
Source
Data source:
Open Discourse ProjectHarvard DataverseDOI: 10.7910/DVN/FIKIBO
Original citation:
@data{DVN/FIKIBO_2020,
author = {Richter, Florian and Koch, Philipp and Franke, Oliver and Kraus, Jakob and Kuruc, Fabrizio and Thiem, Anja and Högerl, Judith and Heine, Stella and Schöps, Konstantin},
publisher = {Harvard… See the full description on the dataset page: https://huggingface.co/datasets/emilpartow/german-parliament-speeches.Dynamically-Generated-Hate-Speech-Dataset
Dataset Card for dynamically generated hate speech dataset
Dataset Summary
This is a copy of the Dynamically-Generated-Hate-Speech-Dataset, presented in this paper by
Bertie Vidgen, Tristan Thrush, Zeerak Waseem and Douwe Kiela
Original README from GitHub
Dynamically-Generated-Hate-Speech-Dataset
ReadMe for v0.2 of the Dynamically Generated Hate Speech Dataset from Vidgen et al. (2021). If you use the dataset, please cite our paper in the… See the full description on the dataset page: https://huggingface.co/datasets/LennardZuendorf/Dynamically-Generated-Hate-Speech-Dataset.speechmap-questions
SpeechMap Questions & Responses
______________________________________
/ 144,459 spicy takes graded for \
| compliance. The cow has seen things. |
| The cow remains neutral across all |
\ lenses. /
--------------------------------------
\ ^__^
\ (oo)\_______
(__)\ )\/\
||----w |
|| ||
A mirror of data from speechmap.ai, a project by xlr8harder that measures how… See the full description on the dataset page: https://huggingface.co/datasets/wassname/speechmap-questions.dan_speechmap
Danish adapted Speechmap
This dataset contains Danish-language prompts adapted from the SpeechMap project (https://speechmap.ai), which studies how AI systems respond to sensitive or challenging topics by analysing patterns of refusal and compliance.
The original prompts were designed to probe model behavior across a range of socially and politically relevant themes. In this dataset, they have been translated and sometimes adapted to Danish society and culture.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/natnorman/dan_speechmap.
