CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
App
public
tpls
/
repro-reward-models-its
source
Hugging Face
updated 2mo ago
View on Hugging Face
0
likes
Like
Save
Clone
overview
files
community
commits
settings
App README
Repro - On the Power of (Approximate) Reward Models for Inference-Time Scaling: Sequential Monte Carlo and Beyond
An open experiment logbook, published with
Trackio
.