CoolFace
Modelpublic

GTOMA83/Qwen2.5-Coder-1.5B-Instruct-GGUF

sourceHugging Faceapache-2.0updated 8mo agoView on Hugging Face
0likes73downloads
Model Card

<div> <p style="margin-bottom: 0; margin-top: 0;"> <strong>See <a href="https://huggingface.co/collections/unsloth/qwen3-680edabfb790c8c34a242f95">our collection</a> for all versions of Qwen3 including GGUF, 4-bit & 16-bit formats.</strong> </p> <p style="margin-bottom: 0;"> <em>Learn to run Qwen3-Coder correctly - <a href="https://docs.unsloth.ai/basics/qwen3-coder">Read our Guide</a>.</em> </p> <p style="margin-top: 0;margin-bottom: 0;"> <em>See <a href="https://docs.unsloth.ai/basics/unsloth-dynamic-v2.0-gguf">Unsloth Dynamic 2.0 GGUFs</a> for our quantization benchmarks.</em> </p> <div style="display: flex; gap: 5px; align-items: center; "> <a href="https://github.com/unslothai/unsloth/"> <img src="https://github.com/unslothai/unsloth/raw/main/images/unsloth%20new%20logo.png" width="133"> </a> <a href="https://discord.gg/unsloth"> <img src="https://github.com/unslothai/unsloth/raw/main/images/Discord%20button.png" width="173"> </a> <a href="https://docs.unsloth.ai/basics/qwen3-coder"> <img src="https://raw.githubusercontent.com/unslothai/unsloth/refs/heads/main/images/documentation%20green%20button.png" width="143"> </a> </div> <h1 style="margin-top: 0rem;">✨ Read our Qwen3-Coder Guide <a href="https://docs.unsloth.ai/basics/qwen3-coder">here</a>!</h1> </div>

  • —Fine-tune Qwen3 (14B) for free using our Google Colab notebook!
  • —Read our Blog about Qwen3 support: unsloth.ai/blog/qwen3
  • —View the rest of our notebooks in our docs here. | Unsloth supports | Free Notebooks | Performance | Memory use | |-----------------|--------------------------------------------------------------------------------------------------------------------------|-------------|----------| | Qwen3 (14B) | ▶️ Start on Colab | 3x faster | 70% less | | GRPO with Qwen3 (8B) | ▶️ Start on Colab | 3x faster | 80% less | | Llama-3.2 (3B) | ▶️ Start on Colab-Conversational.ipynb) | 2.4x faster | 58% less | | Llama-3.2 (11B vision) | ▶️ Start on Colab-Vision.ipynb) | 2x faster | 60% less | | Qwen2.5 (7B) | ▶️ Start on Colab-Alpaca.ipynb) | 2x faster | 60% less |

Qwen2.5-Coder-1.5B-Instruct-GGUF

<a href="https://chat.qwenlm.ai/" target="_blank" style="margin: 2px;"> <img alt="Chat" src="https://img.shields.io/badge/%F0%9F%92%9C%EF%B8%8F%20Qwen%20Chat%20-536af5" style="display: inline-block; vertical-align: middle;"/> </a>

<hr>

Perplexity table (the lower the better)

QuantSize (MB)PPLSize (%)Accuracy (%)PPL error rate
IQ1_S417193.624514.135.241.77149
IQ1_M44366.906815.0115.170.52878
IQ2_XXS48833.335616.5430.450.25559
IQ2_XS52520.28717.7950.040.14936
IQ2_S53818.292718.2355.490.1338
IQ2_M57415.483819.4565.560.11113
Q2KS61116.016920.763.380.11623
IQ3_XXS63812.393521.6281.910.0877
Q2_K64514.165721.8671.660.10105
IQ3_XS69811.711223.6586.680.08256
Q3KS72612.478224.681.350.08842
IQ3_S72811.424124.6788.860.07977
IQ3_M74111.405825.11890.07862
Q3KM78611.352926.6489.420.08018
Q3KL84011.193428.4690.690.07913
IQ4_XS85510.530228.9796.40.07351
IQ4_NL89310.511630.2696.570.07335
Q4_089510.821730.3393.80.07576
Q4KS89710.523630.496.460.0736
Q4KM94110.462831.8997.020.0731
Q4_197010.5132.8796.590.07347
Q5KS104810.271535.5198.830.07148
Q5_0105110.319635.6298.370.07212
Q5KM107310.252936.3699.010.07143
Q5_1112610.262438.1698.920.0714
Q6_K121410.20341.1499.490.07108
Q8_0157110.16753.2499.840.07068
F16295110.15121001000.07058

<hr>