Sudnya/classic-eda-c-trajectories
nano-rl trajectories Agent trajectories from nano-rl: an LLM is asked to specify, test and then implement a small C program, which is compiled and executed inside a confined sandbox and scored against tests the model wrote before it saw its own program. When the program fails, the model is shown the build log and the failing cases and asked to repair it, for up to 200 rounds. Every model turn is one row, including the ones that went nowhere. This is a partial snapshot. 130 of 1… See the full description on the dataset page: https://huggingface.co/datasets/Sudnya/classic-eda-c-trajectories.
065
No card is published for this repository, or it could not be fetched from Hugging Face right now.
