CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
2 results
clinical-evaluation
clinical-evaluation
Search
in
all
models
datasets
apps
agents
people
projects
Datasets
All datasets matching “clinical-evaluation”
ClarusC64 /
clinical-mechanistic-parsimony-evaluation-v0.1
What this dataset tests Whether a model can rank diagnoses by mechanistic parsimonyby counting the extra assumptions needed to fit all evidence. Required outputs parsimony_rank extra_assumptions_count mechanism_stability_flag Mechanism stability flags stable_mechanism patchwork_mechanism unstable_mechanism Typical failures ranking by prevalence instead of mechanistic fit omitting the assumption list calling a patchwork explanation "stable" Suggested prompt wrapper… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-mechanistic-parsimony-evaluation-v0.1.
tabular
text-classification
n<1K
0 likes
12 downloads
8mo ago
Hugging Face
Apps
All apps matching “clinical-evaluation”
Mahrufa /
clinical-rag-evaluation-framework
docker
0 likes
3mo ago
Hugging Face