CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
Model
public
xinlai
/
DeepSeekMath-RL-Step-DPO
source
Hugging Face
apache-2.0
updated 2y ago
View on Hugging Face
2
likes
35
downloads
Like
Save
Clone
overview
files
community
commits
settings
root
file
.gitattributes
2 KB
download
file
config.json
751 B
download
file
generation_config.json
143 B
download
file
model-00001-of-00003.safetensors
4.64 GB
download
file
model-00002-of-00003.safetensors
4.64 GB
download
file
model-00003-of-00003.safetensors
3.59 GB
download
file
model.safetensors.index.json
22 KB
download
file
README.md
1020 B
download
file
special_tokens_map.json
369 B
download
file
tokenizer_config.json
1 KB
download
file
tokenizer.json
4.4 MB
download
file
trainer_state.json
444 KB
download
file
training_args.bin
7 KB
download
xinlai/DeepSeekMath-RL-Step-DPO · main · files are served by the source, never re-hosted here