CoolFace
Datasetpublicgated

Sudnya/classic-eda-c-trajectories

nano-rl trajectories Agent trajectories from nano-rl: an LLM is asked to specify, test and then implement a small C program, which is compiled and executed inside a confined sandbox and scored against tests the model wrote before it saw its own program. When the program fails, the model is shown the build log and the failing cases and asked to repair it, for up to 200 rounds. Every model turn is one row, including the ones that went nowhere. This is a partial snapshot. 279 of 1… See the full description on the dataset page: https://huggingface.co/datasets/Sudnya/classic-eda-c-trajectories.

sourceHugging Faceapache-2.0updated 1d agoView on Hugging Face
0likes82downloads

No commit history came back for main. The revision may not exist, or the source declined the request.