positron-ai/meta-llama_Llama-3.2-1B-Instruct-tron-best-gptq-permuted
Positron AI Quantized Build
Built with Llama.
This repository contains a Positron AI quantized build of meta-llama/Llama-3.2-1B-Instruct for inference.
Recommended Use
Use this artifact when you need a GPTQ 4-bit build of meta-llama/Llama-3.2-1B-Instruct built by Positron AI.
For general-purpose GPU inference, compare against the original model and other quantized formats before deployment.
Artifact Summary
Quantization Details
License
This model is a derivative of Llama 3.2 and is distributed under the Llama 3.2 Community License (LICENSE.txt). The repository also includes Meta's Acceptable Use Policy (USE_POLICY.md) and the attribution notice required by the license (NOTICE).
Provenance
This artifact was produced by Positron AI from meta-llama/Llama-3.2-1B-Instruct via GPTQ quantization using GPTQModel. The original model license and usage restrictions continue to apply.
