models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Devstral-2-123B-NVFP4-TensorRT-LLMCodeLlama-13B-Instruct-GPTQ-TensorRT-LLM-RTX-4090Qwen3-1.7B-TensorRT-LLM-Checkpoint-NVFP4Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-NVFP4Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-BF16Qwen3-4B-TensorRT-LLM-Checkpoint-NVFP4Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-FP8Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-FP16Phi-4-mini-instruct-TensorRT-LLM-Checkpoint-NVFP4Qwen3-0.6B-TensorRT-LLM-Checkpoint-FP16Qwen3-0.6B-TensorRT-LLM-Checkpoint-BF16Qwen3-0.6B-TensorRT-LLM-Checkpoint-FP8Qwen3-1.7B-TensorRT-LLM-Checkpoint-BF16Qwen3-4B-TensorRT-LLM-Checkpoint-FP8Qwen3-0.6B-TensorRT-LLM-Checkpoint-NVFP4Qwen3-1.7B-TensorRT-LLM-Checkpoint-FP16TensorRT-LLM-engine-Mistral-7B-instruct-v0.2Phi-4-mini-instruct-TensorRT-LLM-Checkpoint-FP8Qwen3-1.7B-TensorRT-LLM-Checkpoint-FP8llama-3-8b-tensorrt-llm-tp4llama-3-8b-tensorrt-llm-tp1llama-3-8b-tensorrt-llm-int4-awqtensorrtllm_mixtral_dpo_enginewhisper-turbo-v3-TensorRT-LLMtensorrt-llm-wheelsrtx-5080-tensorrt-llm-0.21.0-whisper-large-v2TensorRT-LLM-IPC-RCE-PoCtensorrt-llm-1.3.0rc15TensorRT-LLM-Windows-RTX40
