CoolFace
20 results

disagreement

avewright /chess-soft-100m-disagreements avewright/chess-soft-100m-disagreements Positions where the greedy policy of avewright/chess-transformer-100m-squares64 disagrees with a strong teacher best move (move_idx). Current upload: 1,782,505 rows in 90 shards. Mix split rows shards labels data/shard_*.parquet 1,768,622 83 teacher MultiPV from avewright/chess-soft-multipv-lichess data/sf19/*.parquet 13,883 7 Stockfish 19 max-Elo MultiPV from new games Stream rows are an argmax filter of… See the full description on the dataset page: https://huggingface.co/datasets/avewright/chess-soft-100m-disagreements.tabularother1M<n<10M0 likes140 downloads19d agoHugging FaceDavidYor06 /llm-disagreement Lenz Frontier-LLM Disagreement — v1.1 Five frontier language models each rated the same 997 real fact-check claims submitted by users of Lenz. This dataset is the per-claim record of where they agreed and where they did not. In 63% of real-world fact-checks, top AI models don't agree on the answer — at least one model dissents from the majority, or no majority forms at all (95% CI 60–66%). At a glance Claims 997 complete (of 1,000 harvested) Models… See the full description on the dataset page: https://huggingface.co/datasets/DavidYor06/llm-disagreement.tabulartext-classification1K<n<10K1 likes109 downloads10d agoHugging FaceRuyuanWan /SBIC_DisagreementThis dataset is processed version of Social Bias Inference Corpus(SBIC) dataset including text, annotator's demographics and the annotation disagreement labels. Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement tabulartext-classification10K<n<100K0 likes50 downloads4y agoHugging FaceRuyuanWan /Dynasent_DisagreementThis dataset is processed version of Dynamic Sentiment Analysis (DynaSent) dataset including text and the annotation disagreement labels. Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement Source Data: Dynamic Sentiment Analysis Dataset(Potts et al. 2021) tabulartext-classification100K<n<1M0 likes41 downloads4y agoHugging FaceRuyuanWan /SChem_DisagreementThis dataset is processed version of Social Chemistry 101(SChem) dataset including text and the annotation disagreement labels. Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement Source Data: Social Chemistry 101(Forbes et al. 2020) tabulartext-classification10K<n<100K0 likes33 downloads4y agoHugging FaceRuyuanWan /Politeness_DisagreementThis dataset is processed version of Stanford Politeness Corpus (Wikipedia) including text and the annotation disagreement labels. Paper: Everyone's Voice Matters: Quantifying Annotation Disagreement Using Demographic Information Authors: Ruyuan Wan, Jaehyung Kim, Dongyeop Kang Github repo: https://github.com/minnesotanlp/Quantifying-Annotation-Disagreement Source Data: Wikipedia Politeness Corpus(Danescu-Niculescu-Mizil et al. 2013) tabulartext-classification1K<n<10K1 likes31 downloads4y agoHugging Face