CoolFace
Modelpublic

CALISTA-INDUSTRY/DeepSeek-R1-Distill-Qwen-1.5B-FineTune

sourceHugging Facemitupdated 2y agoView on Hugging Face
0likes32downloads
Model Card

DeepSeek-R1 Release __________________________________________________________________________________________

โšก Performance on par with OpenAI-o1

๐Ÿ“– Fully open-source model & technical report

๐Ÿ† MIT licensed: Distill & commercialize freely!

๐ŸŒ Website & API are live now! Try DeepThink at chat.deepseek.com today! __________________________________________________________________________________________

๐Ÿ”ฅ Bonus: Open-Source Distilled Models!

๐Ÿ”ฌ Distilled from DeepSeek-R1, 6 small models fully open-sourced

๐Ÿ“ 32B & 70B models on par with OpenAI-o1-mini

๐Ÿค Empowering the open-source community

๐ŸŒ Pushing the boundaries of open AI! _____________________________________________________________________

๐Ÿ› ๏ธ DeepSeek-R1: Technical Highlights

๐Ÿ“ˆ Large-scale RL in post-training

๐Ÿ† Significant performance boost with minimal labeled data

๐Ÿ”ข Math, code, and reasoning tasks on par with OpenAI-o1

๐Ÿ“„ More details: https://github.com/deepseek-ai/DeepSeek-R1/blob/main/DeepSeekR1.pdf ____________________________________________________________________

๐ŸŒ API Access & Pricing

โš™๏ธ Use DeepSeek-R1 by setting model=deepseek-reasoner

๐Ÿ’ฐ $0.14 / million input tokens (cache hit)

๐Ÿ’ฐ $0.55 / million input tokens (cache miss)

๐Ÿ’ฐ $2.19 / million output tokens

๐Ÿ“– API guide: https://api-docs.deepseek.com/guides/reasoning_model