models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
granite-3.1-2b-instruct-quantized.w4a16Meta-Llama-3.1-8B-Instruct-quantized.w4a16Qwen3-8B-quantized.w4a16Qwen3-0.6B-quantized.w4a16Meta-Llama-3.1-8B-Instruct-quantized.w8a8DeepSeek-R1-Distill-Qwen-32B-quantized.w4a16Qwen3-32B-quantized.w4a16Meta-Llama-3.1-70B-Instruct-quantized.w4a16Qwen3-4B-quantized.w4a16Qwen3-30B-A3B-Instruct-2507-quantized.w4a16Qwen3-30B-A3B-Instruct-2507-quantized.w8a8Qwen3-30B-A3B-quantized.w4a16MiniCPM5-2B-RotSVDMix-Quantizedgranite-4.1-3b-quantized.w8a8Phi-3-medium-128k-instruct-quantized.w4a16Qwen3-4B-Instruct-2507-quantized.w8a8Meta-Llama-3.1-70B-Instruct-quantized.w8a8Meta-Llama-3.1-8B-Instruct-quantized.w8a16Mistral-7B-Instruct-v0.3-quantized.w4a16Qwen3-1.7B-quantized.w4a16Qwen3-4B-Instruct-2507-quantized.w4a16Meta-Llama-3.1-8B-quantized.w8a8DeepSeek-R1-Distill-Llama-70B-quantized.w4a16OpenCodeReasoning-Nemotron-1.1-32B-GGUFquantized-gemma-7b-itOpenCodeReasoning-Nemotron-1.1-7B-GGUFgemma-7b-it-GGUF-quantizedQwen3-Next-80B-A3B-Instruct-quantized.w4a16Meta-Llama-3-8B-Instruct-quantized.w8a8OpenCodeReasoning-Nemotron-1.1-14B-GGUF
