models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Meta-Llama-3.1-70B-Instruct-FP8Meta-Llama-3.1-8B-Instruct-FP8Meta-Llama-3.1-8B-Instruct-quantized.w4a16Meta-Llama-3.1-8B-Instruct-FP8-dynamicMeta-Llama-3.1-8B-Instruct-quantized.w8a8Meta-Llama-3.1-70B-Instruct-quantized.w4a16Meta-Llama-3.1-8B-FP8Meta-Llama-3.1-8B-quantized.w8a8Meta-Llama-3.1-405B-Instruct-FP8-dynamicNVIDIA-Nemotron-3-Ultra-550B-A55B-FP8-dynamicNVIDIA-Nemotron-Nano-9B-v2-FP8-dynamicNVIDIA-Nemotron-3-Nano-30B-A3B-FP8NVIDIA-Nemotron-3-Super-120B-A12B-FP8Meta-Llama-3.1-70B-Instruct-quantized.w8a8Meta-Llama-3.1-8B-Instruct-quantized.w8a16Meta-Llama-3.1-405B-Instruct-FP8granite-4.2-3bMeta-Llama-3.1-70B-Instruct-FP8-dynamicNVIDIA-Nemotron-Nano-9B-v2-quantized.w4a16Meta-Llama-3.1-70B-FP8NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4Meta-Llama-3.1-8B-quantized.w8a16Meta-Llama-3.1-405B-FP8Mistral-Medium-3.5-128BMeta-Llama-3.1-70B-Instruct-quantized.w8a16NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16NVIDIA-Nemotron-3-Super-120B-A12B-BF16NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16command-a-plus-05-2026-w4a4
