CoolFace
Datasetpublic

sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation

Reproduction: Off-Policy Learning in Large Action Spaces - Optimization Matters More Than Estimation Paper Information Title: Off-Policy Learning in Large Action Spaces: Optimization Matters More Than Estimation OpenReview ID: srIStBTJiu Conference: ICML 2026 Task: Compare optimization landscapes of IPS vs PWLL for off-policy policy learning Reproduction Summary This reproduction evaluates the paper's core thesis: optimization landscape (not… See the full description on the dataset page: https://huggingface.co/datasets/sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation.

sourceHugging Faceupdated 2mo agoView on Hugging Face
2likes36downloads
2 commits on main
139d7d12mo ago

Add ICML 2026 reproduction logbook for paper srIStBTJiu

sabaridsnfuji
821ee612mo ago

initial commit

sabaridsnfuji