sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation
Reproduction: Off-Policy Learning in Large Action Spaces - Optimization Matters More Than Estimation Paper Information Title: Off-Policy Learning in Large Action Spaces: Optimization Matters More Than Estimation OpenReview ID: srIStBTJiu Conference: ICML 2026 Task: Compare optimization landscapes of IPS vs PWLL for off-policy policy learning Reproduction Summary This reproduction evaluates the paper's core thesis: optimization landscape (not… See the full description on the dataset page: https://huggingface.co/datasets/sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation.
236
