reasoning-degeneration-dev/gepa-rlm-exp-20260219-031221
gepa-rlm-exp-20260219-031221 GEPA vs GEPA+RLM prompt optimization experiment on AIME math problems. Task LM: openai/gpt-4.1-mini | Reflection LM: openai/gpt-5 | Last updated: 2026-02-19 05:07 UTC Results Run Method k Val Score Test Acc Tokens Cost Time fixed_rlm_k20 rlm 20 40.00% 41.33% 1,805,475 $0.0000 5388s Learning Curves Experiment Config { "script_name": "run_experiment.py", "model": "openai/gpt-4.1-mini"… See the full description on the dataset page: https://huggingface.co/datasets/reasoning-degeneration-dev/gepa-rlm-exp-20260219-031221.
045
