models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Qwen3.5-397B-A17B-MXFP4Qwen3.8-2.4T-A95B-Quark-MXFP4Qwen3-VL-8B-Instruct-w8a8-llmcompressorGLM-5.3-Quark-MXFP4-AttnFP8Qwen3.5-397B-A17B-MXFP4-AttnFP8-V2Instella-MoE-16B-A3B-ThinkQwen3.5-397B-A17B-NVFP4gemma-4-12B-it-w4a16-llmcompressorPhi-4-reasoning-plus-w4a16-llmcompressorInstella-MoE-16B-A3B-SFTMiniMax-M2.1-MXFP4whisper-telephony-amdAMD-OLMo-1B-IT-i1-GGUFQwen3.8-Flash-Next-Quark-MXFP4AMD-Llama-135m-code-i1-GGUFQwen3.5-9B-w4a16-asym-torchao-v0.17.0gemma-4-26B-A4B-it-w4a16-llmcompressorMixtral-8x7B-Instruct-v0.1-w8a8-llmcompressor-v0.10.0.2Qwen3.6-35B-A3B-w4a16-llmcompressorLlama-3.1-8B-Instruct-w4a16-asym-torchao-v0.17.0Qwen3.8-Flash-Next-Quark-MXFP4-PLEFP8Mixtral-8x7B-Instruct-v0.1-w4a16-llmcompressorQwen3-VL-8B-Instruct-w4a16-asym-torchao-v0.17.0MiniMax-M2.7-MXFP4Qwen3.6-35B-A3B-w8a8-llmcompressorQwen3-235B-A22B-Instruct-2507-MXFP4AMD-Llama-135m-i1-GGUFPhi-4-reasoning-plus-w4a16-asym-torchao-v0.17.0gpt-oss-20b-BF16-w8a8-llmcompressorLlama3_1-8B-Instruct-AMD-python-i1-GGUF
