satire
Datasets
All datasets matching “satire”ukr-tg-satire
Ukrainian Telegram satire & troll posts
A corpus of 92k posts from 17 public Telegram channels in the
satire/irony/parody register, harvested July 2026 via the public Telegram
API (history reaching back to 2018 for some channels). The channels are
Ukrainian-audience but the posts are a Ukrainian/Russian mix (the
main truha and sria_news feeds write mostly in Russian, the regional
truexa* branches mostly in Ukrainian), so the corpus carries
language: [ru, uk].
Two flavors are… See the full description on the dataset page: https://huggingface.co/datasets/hausmer/ukr-tg-satire.ukr-tg-satire-2
TG channels wave 2
A second scrape wave of 6 public Telegram channels
exported 2026-09-15 from their full histories via a
read-only MTProto API. Same schema as the main corpora
(hausmer/ukr-tg-media,
hausmer/ukr-tg-satire).
24,743 text posts across:
Foma_memes — Мемарня Sa-chan1917|ИзюмТГ (1,892 posts)
zelenskyi_vladimir — Владимир Зеленский (пародия) (612 posts)
karikaturnaya_satira — Карикатурная сатира (138 posts)
memarnya_rezerv — Ініціативна група «20 см» (1,169 posts)… See the full description on the dataset page: https://huggingface.co/datasets/hausmer/ukr-tg-satire-2.truha-news-satire
Truha news satire (Ukrainian)
A synthetic, LLM-generated dataset of Ukrainian "news" written in the
truha satirical register — short, meta-ironic, punchline-driven fake-news
items. The corpus was produced by distilling a target style (a Ukrainian
satirical news persona) into generated examples; no real user data is
included.
Intended use: style-transfer / imitation training for a Ukrainian satirical
news-writing assistant. Each example is a (system, instruction, output) triple.… See the full description on the dataset page: https://huggingface.co/datasets/hausmer/truha-news-satire.Arbic-satire-dataset
Dataset Card for dummy
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/arbml/Arbic-satire-dataset.multimodal_satire
Dataset card for "multimodal_satire"
This is the dataset for the paper A Multi-Modal Method for Satire Detection using Textual and Visual Cues. To obtain the full-text body of the articles, you need to scrape websites using the provided links in the dataset.
GitHub repository: https://github.com/lilyli2004/satire
Reference
If you use this dataset, please cite the following paper:
@inproceedings{li-etal-2020-multi-modal,
title = "A Multi-Modal Method for Satire… See the full description on the dataset page: https://huggingface.co/datasets/phosseini/multimodal_satire.SatireInstruct
Satire Dataset
This is a satire dataset collection, with instruct topics. It is meant to be used in fine-tuning.
More info will be added once the datasets are fully finished.
Folder Structure
Folder
Purpose
data/
The foods and topics datasets used in the pipeline. It may be unused later on in order to make the dataset more consistent.
prompts/
Separate prompt files used for generating the recipes, terrible emails, and main instruct examples.… See the full description on the dataset page: https://huggingface.co/datasets/benni-ben/SatireInstruct.
