datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
task903_deceptive_opinion_spam_classification
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task903_deceptive_opinion_spam_classification
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task903_deceptive_opinion_spam_classification.task902_deceptive_opinion_spam_classification
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task902_deceptive_opinion_spam_classification
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task902_deceptive_opinion_spam_classification.tax-court-opinions
Tax Court Opinions
Text of United States Tax Court opinions, primarily from 1995 through September 2026, with a small number of earlier opinions back to 1986. The Tax Court publishes its opinions as PDF files; these were converted to text using pdfminer.
Dataset Structure
14,848 rows, one per opinion. Columns:
Column
Type
Description
filename
string
Original filename, encoding year/month/day/type/name/pages/docket/judge
year
string
Filing year (four… See the full description on the dataset page: https://huggingface.co/datasets/andrew-mitchel/tax-court-opinions.scotus-opinions-to-2020
Scotus 1970 to 2020 Options
This dataset was constructed based on the Scotus 1970 to 2020 Kaggle data and public opinions found on Courtlistener.
weibo-opinion-dynamic-single-dim
Weibo Sentiment Evolution Dataset
This dataset contains Weibo posts and their associated comment threads used for studying sentiment evolution and opinion dynamics in social media discussions.
The dataset is distributed as a single JSON Lines file:
weibo_dataset.jsonl
Each line is one Weibo post record. Comments for that post are embedded in the comments field.
Dataset Details
Number of post records: 1,379
Number of embedded comments: 93,569
Number of Weibo… See the full description on the dataset page: https://huggingface.co/datasets/hreyulog/weibo-opinion-dynamic-single-dim.opinions
Opinione - Albanian Opinion & Editorial Corpus
Dataset Description
A collection of Albanian-language opinion pieces, editorials, and columns (1,422 summaries) gathered from news portals across Albania, Kosovo, and North Macedonia. It complements the akadriu/lajme-shqip news corpus with subjective/argumentative text for Albanian NLP.
Data Fields
Field
Type
Description
link
string
Original article URL
title
string
Piece headline
portal… See the full description on the dataset page: https://huggingface.co/datasets/akadriu/opinions.nytimes-writing-prompts-student-opinion
Writing prompts - student opinion
Daily questions inspired by Times content from across sections. Join the conversation!
Dataset Details
Source: https://www.nytimes.com/column/learning-student-opinion
Last update: 2025-09-30
Row count: 2,506
Direct Use
text generation
