CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
1 results
llm-failures
llm-failures
Search
in
all
models
datasets
apps
agents
people
projects
Datasets
All datasets matching “llm-failures”
ClarusC64 /
llm-termination-failures
Purpose • Capture termination failures in LLMs • Focus on overcompletion and boundary violations Why it matters • Models often answer correctly • But fail to stop when the task is complete • This failure degrades reliability in real deployments Use cases • Eval benchmarks • Fine tuning stop behavior • Instruction adherence research
0 likes
16 downloads
9mo ago
Hugging Face