CoolFace
Modelpublic

juanquivilla/sotto-cleanup-lfm25-350m

sourceHugging Facemitupdated 5mo agoView on Hugging Face
1likes65downloads
33 commits on main
6df6f015mo ago

soup: corrected README with full prod metrics + soup explanation

juanquivilla
9f2abdb5mo ago

soup_30: composite=89.45 — see model card for benchmark deltas vs v45

juanquivilla
44245045mo ago

v55: corrected metrics with max_new=900 (sub_del15 3.7%, composite 89.32)

juanquivilla
7fbef215mo ago

v55: add inference recommendations (rep_pen=1.05, max_new_tokens guidance)

juanquivilla
d5ba9555mo ago

v55: composite=88.95 — see model card for benchmark deltas vs v45

juanquivilla
e8ff65e5mo ago

v51: composite=88.68 — see model card for benchmark deltas vs v45

juanquivilla
81578ee5mo ago

v45: SFT+chained GRPO with ITN — 95.9% number accuracy, 97.0% filler-free, deletion behavior matches v36

juanquivilla
72782275mo ago

v36: full-FT GRPO with substantive-deletion-aware reward — filler-free 96.9%, sub-del-15-long 0.64%

juanquivilla
15c2adb6mo ago

v23 R6 (4-stage SFT->GRPO->Stage2->GRPO): ROUGE-L 0.9537 (tied v22), Filler-Free 91.1% (beats v22 90.3%), paragraph rate 91.5% — definitive v23 model

juanquivilla
ae896a96mo ago

v23 R5 (paragraph rows excluded from GRPO): ROUGE-L 0.9505, Filler-Free 91.0%, paragraph rate 89.5% — best v23 variant overall

juanquivilla
56ac4306mo ago

v23 R4 (LR 5e-6): Filler-Free 91.0% (beats v22 90.3%), paragraph rate 89%, ROUGE-L 0.9499

juanquivilla
8d24c186mo ago

v23+paragraphs: ROUGE-L 0.9506, Filler-Free 90.2%, paragraph rate 91.5% (0% in v22)

juanquivilla
41110166mo ago

v22+GRPO-r32: ROUGE-L 0.954 val, 66% exact, 91% filler-free — new production model

juanquivilla
5f57a196mo ago

v22+GRPO: ROUGE-L 0.953 val set, 91% filler-free — GRPO works on proper benchmark

juanquivilla
607327b6mo ago

v22-lr3: ROUGE-L 0.948 on val set — cleaned data + LR 3e-5

juanquivilla
69428ac6mo ago

v18: ROUGE-L 0.968, 72% exact — AdamW beta2=0.95 breakthrough

juanquivilla
257897f6mo ago

v17: ROUGE-L 0.966, 71% exact — preserve phrases + redundant tail removal

juanquivilla
00c58076mo ago

v16: ROUGE-L 0.962, preserve-phrase data (crutch_words 0.987)

juanquivilla
b64dd3c6mo ago

Update model card: v15 detailed card with examples, benchmarks, and links

juanquivilla
27c16fc6mo ago

v15: ROUGE-L 0.960, 70% exact match — LR 2.5e-5 breakthrough

juanquivilla
14627236mo ago

v7: combined dataset LR 2e-5, ROUGE-L 0.943, 62% exact

juanquivilla
4dd89e66mo ago

v5: LR 2e-5 + Stage-2, ROUGE-L 0.942, 60% exact — new all-time record

juanquivilla
23ff3b56mo ago

v4: 116K data + long transcripts, ROUGE-L 0.927, handles 1000+ word inputs

juanquivilla
1c9bfcc6mo ago

Upload README.md with huggingface_hub

juanquivilla
9154aef6mo ago

Stage-2: ROUGE-L 0.931, 90% zero-filler — new best

juanquivilla
1811cf86mo ago

Full FT v2: ROUGE-L 0.930, 55% exact — best model

juanquivilla
840b0716mo ago

Full FT + GRPO: ROUGE-L 0.916 — new record, +2.5 over 2B

juanquivilla
1e7e04f6mo ago

Full FT: ROUGE-L 0.907 — new record, +1.6 over prompted 2B

juanquivilla
3ac09996mo ago

GRPO R2: ROUGE-L 0.892 — exceeds prompted 2B

juanquivilla
4ab77086mo ago

GRPO model: ROUGE-L 0.891 — matches prompted 2B

juanquivilla
88c1f586mo ago

Upload folder using huggingface_hub

juanquivilla
eda4f6c6mo ago

Upload README.md with huggingface_hub

juanquivilla
83e83fb6mo ago

initial commit

juanquivilla