CoolFace
Modelpublic

Fazmin/solus_v1_deepseek-r1-distill-qwen-7b-q4

sourceHugging Facemitupdated 2mo agoView on Hugging Face
0likes23downloads
Model Card

DeepSeek R1 Distill Qwen 7B — Solus v1

A reasoning model built by distilling DeepSeek-R1's chain-of-thought behaviour into a Qwen2.5 7B base. It works through problems step by step before committing to an answer, which makes it markedly stronger than similarly sized models on maths, logic puzzles, and multi-step analysis.

The trade-off is verbosity and latency: it spends tokens thinking, so it is slower to first useful output than a plain instruct model of the same size. MIT licensed.

Specifications

Parameters7B
QuantizationQ4KM
File size4.36 GB
Minimum RAM8.00 GB
Minimum VRAM6.00 GB
Context length32,768 tokens
SHA-256731ece8d06dc7eda6f6572997feb9ee1258db0784827e642909d9b565641937b

Single file: DeepSeek-R1-Distill-Qwen-7B-Q4_K_M.gguf

Quantization

Quantization performed at the Faculty of Engineering, McMaster University.

The GGUF conversion this build is derived from was produced by bartowski, and the weights here are a byte-for-byte copy of that file — the SHA-256 above matches the upstream artifact.

Provenance

Usage

bash
llama-cli -m DeepSeek-R1-Distill-Qwen-7B-Q4_K_M.gguf -cnv