CoolFace
Datasetpublic

sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation

Reproduction: Off-Policy Learning in Large Action Spaces - Optimization Matters More Than Estimation Paper Information Title: Off-Policy Learning in Large Action Spaces: Optimization Matters More Than Estimation OpenReview ID: srIStBTJiu Conference: ICML 2026 Task: Compare optimization landscapes of IPS vs PWLL for off-policy policy learning Reproduction Summary This reproduction evaluates the paper's core thesis: optimization landscape (not… See the full description on the dataset page: https://huggingface.co/datasets/sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation.

sourceHugging Faceupdated 2mo agoView on Hugging Face
2likes36downloads

sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation · main · files are served by the source, never re-hosted here