models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Meta-Llama-3.1-8B-Instruct-AWQ-INT4Meta-Llama-3.1-70B-Instruct-AWQ-INT4Meta-Llama-3.3-70B-Instruct-AWQ-INT4Mixtral-8x7B-Instruct-v0.1-AWQ-INT4Meta-Llama-3.1-8B-Instruct-GPTQ-INT4Llama-3.1-8B-Instruct-AWQ-INT4nemotron-3.5-asr-streaming-0.6b-onnx-int4LTX_2.5_INT4_W4A8_ConvRotLFM2.5-8B-A1B-AWQ-INT4Audio8-TTS-Preview-0.6B-ONNX-INT4Meta-Llama-3.1-405B-Instruct-AWQ-INT4Phi-4-mini-instruct-int4-ovwhisper-large-v3-turbo-int4-ovMistral-Medium-3.5-128B-AWQ-INT4HY-MT1.5-1.8B-GPTQ-Int4Meta-Llama-3.3-70B-Instruct-AWQ-INT4Qwen3.8-27B-int4-AutoRound-SARLlama-3.2-1B-Instruct-SpinQuant_INT4_EO8-ETFalcon3-10B-Instruct-GPTQ-Int4Meta-Llama-3.1-405B-Instruct-GPTQ-INT4Llama-3.2-3B-Instruct-AWQ-INT4Meta-Llama-3.1-70B-Instruct-GPTQ-INT4Llama-3.2-1B-Instruct-SpinQuant_INT4_EO8Llama-3.2-1B-Instruct-QLORA_INT4_EO8Qwen3.5-397B-A17B-heretic-int4-AutoRoundLlama-3.2-3B-Instruct-SpinQuant_INT4_EO8Llama-3.2-3B-Instruct-QLORA_INT4_EO8LFM2.5-8B-A1B-int4-ovMistral-Small-24B-Instruct-2501-GPTQ-INT4Nvidia-Llama-3.1-Nemotron-70B-Instruct-HF-AWQ-INT4
