CoolFace
Datasetpublic

reasoning-degeneration-dev/gepa-rlm-exp-domain_heuristics-20260219-191545

gepa-rlm-exp-domain_heuristics-20260219-191545 GEPA prompt optimization experiment on AIME math problems. Task LM: openai/gpt-4.1-mini | Reflection LM: openai/gpt-5 | Reflection Mode: domain_heuristics | Last updated: 2026-02-19 21:44 UTC Results Run Method k Mode Val Score Test Acc Tokens Cost Time fixed_rlm_k20 rlm 20 domain_heuristics 55.56% 28.00% 1,168,868 $0.0000 6304s Learning Curves Experiment Config {… See the full description on the dataset page: https://huggingface.co/datasets/reasoning-degeneration-dev/gepa-rlm-exp-domain_heuristics-20260219-191545.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes58downloads

reasoning-degeneration-dev/gepa-rlm-exp-domain_heuristics-20260219-191545 · main · files are served by the source, never re-hosted here