datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
stability-recovery-geometry-v0.1
What this dataset does
This dataset tests whether a model can evaluate recovery geometry.
The task is simple:
Given a scenario and a recovery-geometry claim, predict whether the claim is supported.
Core stability idea
Recovery is not merely the return of performance.
Recovery geometry evaluates whether the system is moving toward a healthier basin of operation.
Favorable recovery geometry typically includes:
restoration of function
restoration of margin
reduction of… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/stability-recovery-geometry-v0.1.cfir-stability-intervention-geometry-v0.1CFIR v0.1 — Coupled Failure Intervention Reasoning Benchmark
What this repo does
CFIR v0.1 is a synthetic benchmark designed to evaluate whether models can reason about stability and intervention geometry in coupled systems.
Many real-world failures occur not because systems lack information, but because they fail to interpret interacting pressures, buffers, delays, and couplings correctly.
This dataset tests whether a model can determine when an intervention will stabilize or fail to… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/cfir-stability-intervention-geometry-v0.1.
