models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
DeepSeek-V3-int4-TensorRTDevstral-2-123B-NVFP4-TensorRT-LLMCodeLlama-13B-Instruct-GPTQ-TensorRT-LLM-RTX-4090Qwen3-1.7B-TensorRT-LLM-Checkpoint-NVFP4Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-NVFP4Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-BF16Qwen3-4B-TensorRT-LLM-Checkpoint-NVFP4Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-FP8Qwen2.5-1.5B-Instruct-TensorRT-LLM-Checkpoint-FP16Phi-4-mini-instruct-TensorRT-LLM-Checkpoint-NVFP4Qwen3-0.6B-TensorRT-LLM-Checkpoint-FP16bugged-myelin-tensorrt-gptjQwen3-0.6B-TensorRT-LLM-Checkpoint-BF16Qwen3-0.6B-TensorRT-LLM-Checkpoint-FP8Qwen3-1.7B-TensorRT-LLM-Checkpoint-BF16Qwen3-4B-TensorRT-LLM-Checkpoint-FP8Qwen3-0.6B-TensorRT-LLM-Checkpoint-NVFP4Qwen3-1.7B-TensorRT-LLM-Checkpoint-FP16Phi-4-mini-instruct-TensorRT-LLM-Checkpoint-FP8Qwen3-1.7B-TensorRT-LLM-Checkpoint-FP8TensorRT_Llama-3.1-Minitron-4B-Width-Basegpt-j-6B-tensorrt-int8gpt2-tensorrtchatglm3_6b_32k_TensorRTReady
