Sudnya/classic-eda-c-trajectories
nano-rl trajectories Agent trajectories from nano-rl: an LLM is asked to specify, test and then implement a small C program, which is compiled and executed inside a confined sandbox and scored against tests the model wrote before it saw its own program. When the program fails, the model is shown the build log and the failing cases and asked to repair it, for up to 200 rounds. Every model turn is one row, including the ones that went nowhere. This is a partial snapshot. 279 of 1… See the full description on the dataset page: https://huggingface.co/datasets/Sudnya/classic-eda-c-trajectories.
082
No commit history came back for main. The revision may not exist, or the source declined the request.
