ashritha0907/replay-gap-trajectories
The Replay Gap: Branched Agent Trajectories Counterfactual ("branched") agent rollouts for studying per-step model switching in LLM agents, from the paper The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World (Efficient Reasoning Workshop @ COLM 2026). Routing benchmarks score routers by replaying logged model outputs. In a multi-step agent that is unsound: swap the model at step k and the rest of the trajectory diverges. This dataset contains… See the full description on the dataset page: https://huggingface.co/datasets/ashritha0907/replay-gap-trajectories.
0101
Update README.md
Upload 7 files
initial commit
