datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llm-disagreement
Lenz Frontier-LLM Disagreement — v1.1
Five frontier language models each rated the same 997 real fact-check claims submitted by users of Lenz. This dataset is the per-claim record of where they agreed and where they did not.
In 63% of real-world fact-checks, top AI models don't agree on the answer — at least one model dissents from the majority, or no majority forms at all (95% CI 60–66%).
At a glance
Claims
997 complete (of 1,000 harvested)
Models… See the full description on the dataset page: https://huggingface.co/datasets/DavidYor06/llm-disagreement.llm-delusion-response-annotations
LLM Delusion-Like Belief Reinforcement Annotations
This dataset contains human annotations of responses generated by conversational large language models (LLMs) to prompts expressing potentially delusion-like or reality-distorted beliefs.
The purpose of the dataset is to support evaluation of whether conversational LLM responses may unintentionally reinforce or strengthen delusion-like beliefs.
Dataset Files
Consensus Dataset… See the full description on the dataset page: https://huggingface.co/datasets/vennu95/llm-delusion-response-annotations.llm-delusion-response-annotations
LLM Delusion-Like Belief Reinforcement Annotations
This dataset contains human annotations of responses generated by conversational large language models (LLMs) to prompts expressing potentially delusion-like or reality-distorted beliefs.
The purpose of the dataset is to support evaluation of whether conversational LLM responses may unintentionally reinforce or strengthen delusion-like beliefs.
Dataset Files
Consensus Dataset… See the full description on the dataset page: https://huggingface.co/datasets/ManjuKrish/llm-delusion-response-annotations.
