datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mechanistic-interpretability-papers
Mechanistic Interpretability Papers — FineSet
A research-paper dataset on Mechanistic Interpretability Papers, assembled, deduplicated, and quality-scored by
FineSet from arXiv and Semantic Scholar.
📸 This is a dated snapshot — generated 2026-06-12.
It is not auto-updated. Research on Mechanistic Interpretability Papers moves fast — new papers land on arXiv every
week. Want this same dataset refreshed daily, on a topic you choose? See the bottom. ↓
Why this… See the full description on the dataset page: https://huggingface.co/datasets/fineset-io/mechanistic-interpretability-papers.clinical-mechanistic-parsimony-evaluation-v0.1What this dataset tests
Whether a model can rank diagnoses by mechanistic parsimonyby counting the extra assumptions needed to fit all evidence.
Required outputs
parsimony_rank
extra_assumptions_count
mechanism_stability_flag
Mechanism stability flags
stable_mechanism
patchwork_mechanism
unstable_mechanism
Typical failures
ranking by prevalence instead of mechanistic fit
omitting the assumption list
calling a patchwork explanation "stable"
Suggested prompt wrapper… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-mechanistic-parsimony-evaluation-v0.1.
