models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Qwen3.5-397B-A17B-MXFP4Qwen3.8-2.4T-A95B-Quark-MXFP4GLM-5.3-Quark-MXFP4-AttnFP8Qwen3-VL-8B-Instruct-w8a8-llmcompressorQwen3.5-397B-A17B-MXFP4-AttnFP8-V2Instella-MoE-16B-A3B-ThinkQwen3.5-397B-A17B-NVFP4gemma-4-12B-it-w4a16-llmcompressorQwen3.8-Flash-Next-Quark-MXFP4Phi-4-reasoning-plus-w4a16-llmcompressorInstella-MoE-16B-A3B-SFTwhisper-telephony-amdAMD-Llama-135m-code-i1-GGUFMiniMax-M2.1-MXFP4AMD-OLMo-1B-IT-i1-GGUFQwen3.6-35B-A3B-w4a16-llmcompressorgemma-4-26B-A4B-it-w4a16-llmcompressorMixtral-8x7B-Instruct-v0.1-w8a8-llmcompressor-v0.10.0.2Qwen3.5-9B-w4a16-asym-torchao-v0.17.0Llama-3.1-8B-Instruct-w4a16-asym-torchao-v0.17.0AMD-Llama-135m-i1-GGUFQwen3.8-Flash-Next-Quark-MXFP4-PLEFP8Mixtral-8x7B-Instruct-v0.1-w4a16-llmcompressorQwen3-VL-8B-Instruct-w4a16-asym-torchao-v0.17.0Qwen3-235B-A22B-Instruct-2507-MXFP4Phi-4-reasoning-plus-w4a16-asym-torchao-v0.17.0Qwen3.6-35B-A3B-w8a8-llmcompressorMiniMax-M2.7-MXFP4gpt-oss-20b-BF16-w8a8-llmcompressorLlama3_1-8B-Instruct-AMD-python-i1-GGUF
