CoolFace
Modelpublic

mario-rc/emotional-rlaif-ppo-llama-3.2-3b-instruct

sourceHugging Facellama3.2updated 9d agoView on Hugging Face
0likes88downloads
12 commits on main
9ba89799d ago

docs: correct project repository link

mario-rc
8d61e759d ago

Reorder released models and apply requested card layout adjustments

mario-rc
a7fcf8810d ago

Standardize License section while preserving base-model requirements

mario-rc
0a499ec10d ago

Add brief limitations introduction

mario-rc
3bca1fe10d ago

Apply approved model-card format with model-specific data and examples

mario-rc
3fb66e310d ago

Restore original model-card text with explicitly requested updates

mario-rc
779d4d011d ago

Record Hub-generated LFS attributes in release checksums

mario-rc
b553e2911d ago

Preserve exact frozen reward adapter saved with PPO checkpoint

mario-rc
b701c5c11d ago

Publish verified emotional RLAIF adapter and unified 20-model card

mario-rc
1dfb7d23mo ago

Normalize model card tables

mario-rc
c47fd353mo ago

Upload updated SML emotional RLAIF adapter

mario-rc
d58e5ac3mo ago

initial commit

mario-rc