disagreement
qwen3-4b-constructive-disagreement-loraassertion_sentence_has_disagreement_or_challengedeepseek-r1-disagreementPoliteness_RoBERTa_Text_Disagreement_PredictorSBIC_RoBERTa_Text_Disagreement_PredictorSChem_RoBERTa_Text_Disagreement_PredictorSBIC_RoBERTa_Text_Disagreement_Binary_ClassifierSChem_RoBERTa_Text_Disagreement_Binary_Classifier
chess-soft-100m-disagreements
avewright/chess-soft-100m-disagreements
Positions where the greedy policy of
avewright/chess-transformer-100m-squares64
disagrees with a strong teacher best move (move_idx).
Current upload: 1,782,505 rows in 90 shards.
Mix
split
rows
shards
labels
data/shard_*.parquet
1,768,622
83
teacher MultiPV from avewright/chess-soft-multipv-lichess
data/sf19/*.parquet
13,883
7
Stockfish 19 max-Elo MultiPV from new games
Stream rows are an argmax filter of… See the full description on the dataset page: https://huggingface.co/datasets/avewright/chess-soft-100m-disagreements.llm-disagreement
Lenz Frontier-LLM Disagreement — v1.1
Five frontier language models each rated the same 997 real fact-check claims submitted by users of Lenz. This dataset is the per-claim record of where they agreed and where they did not.
In 63% of real-world fact-checks, top AI models don't agree on the answer — at least one model dissents from the majority, or no majority forms at all (95% CI 60–66%).
At a glance
Claims
997 complete (of 1,000 harvested)
Models… See the full description on the dataset page: https://huggingface.co/datasets/DavidYor06/llm-disagreement.SBIC_DisagreementThis dataset is processed version of Social Bias Inference Corpus(SBIC) dataset including text, annotator's demographics and the annotation disagreement labels.
Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information
Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang
Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement
Dynasent_DisagreementThis dataset is processed version of Dynamic Sentiment Analysis (DynaSent) dataset including text and the annotation disagreement labels.
Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information
Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang
Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement
Source Data: Dynamic Sentiment Analysis Dataset(Potts et al. 2021)
SChem_DisagreementThis dataset is processed version of Social Chemistry 101(SChem) dataset including text and the annotation disagreement labels.
Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information
Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang
Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement
Source Data: Social Chemistry 101(Forbes et al. 2020)
Politeness_DisagreementThis dataset is processed version of Stanford Politeness Corpus (Wikipedia) including text and the annotation disagreement labels.
Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information
Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang
Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement
Source Data: Wikipedia Politeness Corpus(Danescu-Niculescu-Mizil et al. 2013)
