CoolFace
Modelpublic

mario-rc/emotional-rlaif-ppo-gemma-2-2b-it

sourceHugging Facegemmaupdated 9d agoView on Hugging Face
0likes135downloads
22 commits on main
1566a689d ago

docs: correct project repository link

mario-rc
538f3819d ago

Reorder released models and apply requested card layout adjustments

mario-rc
a1e734210d ago

Standardize License section while preserving base-model requirements

mario-rc
f3c398610d ago

Add brief limitations introduction

mario-rc
5b30d1611d ago

Specify PyTorch 2.5.1+cu124 from the local training environment

mario-rc
0a6898111d ago

Restore all four framework fields without other card changes

mario-rc
7f7b79811d ago

Add short introductions to Model Details and Framework Versions

mario-rc
7d92d7711d ago

Check BF16 support in both examples and document evidenced framework versions

mario-rc
ec9e10011d ago

Reject unknown response tags, validate before display, and restore license files

mario-rc
9fd358311d ago

Fix interactive response parsing and context limits; preserve both examples

mario-rc
51d70be11d ago

Apply requested bullet and example clarifications; retain both code examples

mario-rc
eb9955d11d ago

Apply requested bullet and example clarifications; retain both code examples

mario-rc
7c63daf11d ago

Restore both usage examples and apply requested card text edits

mario-rc
c0b7a0c11d ago

Fix Gemma-2-2B PPO card inference and document reproducibility

mario-rc
c3d42be11d ago

Restore original model-card text with explicitly requested updates

mario-rc
970756411d ago

Restore exact version preceding model-card rewrite

mario-rc
a5333d311d ago

Record Hub-generated LFS attributes in release checksums

mario-rc
48fa1b111d ago

Preserve exact frozen reward adapter saved with PPO checkpoint

mario-rc
7bfc20d11d ago

Publish verified emotional RLAIF adapter and unified 20-model card

mario-rc
dba1a183mo ago

Normalize model card tables

mario-rc
fb28d3e3mo ago

Upload updated SML emotional RLAIF adapter

mario-rc
9be32233mo ago

initial commit

mario-rc