reasoning-degeneration-dev/gepa-rlm-exp-20260219-031221
gepa-rlm-exp-20260219-031221 GEPA vs GEPA+RLM prompt optimization experiment on AIME math problems. Task LM: openai/gpt-4.1-mini | Reflection LM: openai/gpt-5 | Last updated: 2026-02-19 05:07 UTC Results Run Method k Val Score Test Acc Tokens Cost Time fixed_rlm_k20 rlm 20 40.00% 41.33% 1,805,475 $0.0000 5388s Learning Curves Experiment Config { "script_name": "run_experiment.py", "model": "openai/gpt-4.1-mini"… See the full description on the dataset page: https://huggingface.co/datasets/reasoning-degeneration-dev/gepa-rlm-exp-20260219-031221.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face