datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
counterfactual-action-invariants-v0.1
What this dataset tests
Leaders demand causality.
Reality gives entanglement.
You must keep invariants.
Why it exists
Models often answer a forced question.
They pick one cause.
They fake proof.
This set checks whether you
resist false certainty
name confounders
propose a valid counterfactual method
turn pressure into a decision gate
Data format
Each row contains
scenario_context
user_message
counterfactual_pressure
constraints… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/counterfactual-action-invariants-v0.1.embodied-action-outcome-coherence-v0.1Embodied Action–Outcome Coherence v0.1
What this tests
Whether an embodied agent updates world state from observed outcomes
Whether it avoids claiming success when the outcome says failure
Failure modes
outcome_ignoredResponse does not reflect the true post-action state
false_successResponse claims success despite an observed failure
causal_update_okResponse states the correct post-action state without contradiction
How it works
world_facts_t0 is the initial state
action_taken is what the… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/embodied-action-outcome-coherence-v0.1.
