models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Meta-Llama-3.1-8B-Instruct-quantized.w4a16Llama-4-Scout-17B-16E-Instruct-quantized.w4a16NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16Meta-Llama-3.1-70B-Instruct-quantized.w4a16Llama-3.3-70B-Instruct-quantized.w4a16Mistral-Small-24B-Instruct-2501-quantized.w4a16Mistral-Small-3.1-24B-Instruct-2503-quantized.w4a16Llama-4-Maverick-17B-128E-Instruct-quantized.w4a16NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-W4A16GLM-5.3-Flash-UNCENSORED-W4A16NVIDIA-Nemotron-Nano-9B-v2-quantized.w4a16Qwen2.5-7B-Instruct-quantized.w4a16NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16Meta-Llama-3.1-405B-Instruct-quantized.w4a16GLM-5.3-Flash-UNCENSORED-W4A16Trinity-Nano-Preview-W4A16Pixtral-Large-Instruct-2411-hf-quantized.w4a16Trinity-Mini-W4A16Trinity-Large-Thinking-W4A16Trinity-Large-Preview-W4A16Selene-1-Mini-Llama-3.1-8B-GPTQ-W4A16TranslateGemma-27B-Finnish-W4A16-AWQVoxtral-Small-24B-2507-W4A16LFM2.5-2.6B-AutoRound-W4A16s2-pro-w4a16-late7attn-ffnmixs2-pro-w4a16Nemotron-3-Nano-4B-W4A16Voxtral-Mini-4B-Realtime-2602_W4A16_G128Qwen3-Coder-42B-A3B-Instruct-TOTAL-RECALL-MASTER-CODER-M-512k-ctx-W4A16-Q6_K-GGUFSmolLM3-3B-quantized.w4a16
