CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
Model
public
RegularizedSelfPlay
/
Llama-3-8B-Instruct-SPPO-Iter1-gp-8b-gpm-reg0.05-sppo-reversekl-table
source
Hugging Face
updated 1y ago
View on Hugging Face
0
likes
16
downloads
Like
Save
Clone
overview
files
community
commits
settings
3 commits on main
7f23fd7
1y ago
Upload tokenizer
timxiaohangt
1aeee0b
1y ago
Upload LlamaForCausalLM
timxiaohangt
60e5d74
1y ago
initial commit
timxiaohangt