models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Meta-Llama-3.1-8B-Instruct-quantized.w4a16Meta-Llama-3.1-8B-Instruct-quantized.w8a8Meta-Llama-3.1-70B-Instruct-quantized.w4a16Llama-3.2-1B-Instruct-quantized.w8a8Llama-3.3-70B-Instruct-quantized.w4a16Mistral-Small-24B-Instruct-2501-quantized.w4a16Mistral-Small-24B-Instruct-2501-quantized.w8a8Qwen2.5-7B-Instruct-quantized.w8a8Meta-Llama-3.1-70B-Instruct-quantized.w8a8Meta-Llama-3.1-8B-Instruct-quantized.w8a16Llama-3.3-70B-Instruct-quantized.w8a8Llama-3.2-3B-Instruct-quantized.w8a8Meta-Llama-3.1-8B-quantized.w8a8Qwen2.5-7B-Instruct-quantized.w4a16NVIDIA-Nemotron-Nano-9B-v2-quantized.w4a16Meta-Llama-3.1-405B-Instruct-quantized.w4a16Llama-3.2-3B-Instruct-quantized.w8a8Meta-Llama-3.1-8B-quantized.w8a16Meta-Llama-3.1-70B-Instruct-quantized.w8a16Meta-Llama-3.1-405B-Instruct-quantized.w8a16Meta-Llama-3.1-405B-Instruct-quantized.w8a8llama-3.2-1b-instruct-mlx-quantizedaya-23-8B-quantizedLlama-3.1-8b-instruct-quantizedministral-3-3B-it2512-mlx-quantizedggml-LLaMa-65B-quantizedLing-mini-2.0-QuantizedLFM2.5-1.2B-Base-Quantized
