CoolFace
Modelpublic

mario-rc/emotional-rlaif-dpo-gemma-2-2b-it

sourceHugging Facegemmaupdated 9d agoView on Hugging Face
0likes109downloads
50 commits on main
6d2ca549d ago

docs: correct project repository link

mario-rc
457fbad9d ago

Reorder released models and apply requested card layout adjustments

mario-rc
35f006d10d ago

Standardize License section while preserving base-model requirements

mario-rc
4a2ec6310d ago

Add brief limitations introduction

mario-rc
cec7bed11d ago

Restore all four framework fields without other card changes

mario-rc
0bf60f711d ago

Add short introductions to Model Details and Framework Versions

mario-rc
232d83a11d ago

Check BF16 support in both examples and document evidenced framework versions

mario-rc
714754d11d ago

Reject unknown response tags, validate before display, and restore license files

mario-rc
cbca5f511d ago

Fix interactive response parsing and context limits; preserve both examples

mario-rc
e7c95a311d ago

Apply requested bullet and example clarifications; retain both code examples

mario-rc
2ec538c11d ago

Apply requested bullet and example clarifications; retain both code examples

mario-rc
19691f411d ago

Restore both usage examples and apply requested card text edits

mario-rc
948dc2e11d ago

Apply minimal inference and reproducibility fixes to Gemma-2-2B DPO card

mario-rc
84b850011d ago

Restore original model-card text with explicitly requested updates

mario-rc
de5655e11d ago

Restore exact version preceding model-card rewrite

mario-rc
6146dff11d ago

Record Hub-generated LFS attributes in release checksums

mario-rc
5c6ec9911d ago

Publish verified emotional RLAIF adapter and unified 20-model card

mario-rc
f2bb03a3mo ago

Normalize model card tables

mario-rc
89d21793mo ago

Upload updated SML emotional RLAIF adapter

mario-rc
25a55114mo ago

Reorder model comparison tables

mario-rc
403c6c94mo ago

Polish training data paragraph

mario-rc
cda92bc4mo ago

Fix remote code flags

mario-rc
c9d44d94mo ago

Fix DPO data description

mario-rc
91cf55b4mo ago

Adapt interactive prompt templates

mario-rc
6fb65074mo ago

Clarify training data sources

mario-rc
1bbc0b84mo ago

Normalize training data paragraph

mario-rc
47a4f704mo ago

Update usage examples

mario-rc
d94275c4mo ago

Fix training data description

mario-rc
6a38cee4mo ago

Refine model card

mario-rc
1460db44mo ago

Move training before evaluation and simplify metrics

mario-rc
f8e148c4mo ago

Use one evaluation table for released models

mario-rc
589cbd04mo ago

Add evaluation section

mario-rc
e3032ec4mo ago

Add project repository link

mario-rc
1bfe0624mo ago

Add released emotional RLAIF model comparison table

mario-rc
417bc894mo ago

Update model card title

mario-rc
11864f64mo ago

Update model card after repository rename

mario-rc
33cbc5f4mo ago

Format dataset subset names in introduction

mario-rc
0f8a95e4mo ago

Standardize model details and training format

mario-rc
38135f14mo ago

Clarify dataset usage in model card introduction

mario-rc
047e97b4mo ago

Simplify training procedure section

mario-rc
ece69b84mo ago

Link dataset in model card introduction

mario-rc
aecc2c04mo ago

Standardize training and framework sections

mario-rc
35f6b3a4mo ago

Link training dataset in model card

mario-rc
e430d834mo ago

Bold model card key-value labels

mario-rc
e6b05cb4mo ago

Normalize model card section order

mario-rc
0b79bf84mo ago

Place intended use and limitations sections

mario-rc
24ac84f4mo ago

Place framework versions after training procedure

mario-rc
7489ec14mo ago

Add simple and interactive usage examples

mario-rc
c94d59a4mo ago

Reorder usage section and remove evaluation from model card

mario-rc
46af5fd4mo ago

Append long usage example to model card

mario-rc