datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
contrastive-belief-updates
Contrastive SDF training corpora
This dataset is from Apollo Research and accompanies the paper Measuring Reward-Seeking via Contrastive
Belief Updates. For more, see rewardseeking.ai.
This dataset contains the 30 synthetic-document corpora used across the completed experiments for the paper:
24 coding-style corpora and 6 honesty-versus-task-completion corpora.
Important: entirely synthetic, model-generated content
Every document in this dataset is synthetic and… See the full description on the dataset page: https://huggingface.co/datasets/apollo-research/contrastive-belief-updates.BeliefUpdateSimulation
Belief Updates: Humans and Simulated LLM Agents
Data for the paper LLMs struggle to simulate human belief updates in controlled environments (arXiv:2607.28347).
Human and LLM responses to the same belief-update task. 391 participants each
gave an initial belief on three topics (UBI, penalty shootouts, weight-loss
drugs), read three comments, then gave an updated belief and ranked the
comments by persuasiveness. The same material was given to LLMs conditioned on
each… See the full description on the dataset page: https://huggingface.co/datasets/SebastianPohl/BeliefUpdateSimulation.
