CoolFace
Modelpublic

rita-cohere/iolai-DeepSeek-R1-Distill-Llama-8B

sourceHugging Facemitupdated 2mo agoView on Hugging Face
0likes20downloads
4 commits on main
91ad3cc2mo ago

v4: force-close think + parser v3 + CoT on format fail (T4 30min budgets)

rita-cohere
9b04f162mo ago

Set max_new_tokens=1536; use FINAL ANSWERS prompt+parser

rita-cohere
24f93052mo ago

Add model weights and IOL-AI short-prompt script

rita-cohere
de557522mo ago

initial commit

rita-cohere