positron-ai/google_gemma-4-12B-it-ingest-best-gptq-permuted
0680
Positron AI Quantized Build
This repository contains a Positron AI quantized build of google/gemma-4-12B-it for inference.
Recommended Use
Use this artifact when you need a GPTQ 4-bit build of google/gemma-4-12B-it built by Positron AI.
For general-purpose GPU inference, compare against the original model and other quantized formats before deployment.
Artifact Summary
Quantization Details
Evaluation
This card intentionally reports no performance or quality metrics (no KL-divergence, accuracy, or perplexity figures). Validation results are tracked internally by Positron AI.
Provenance
This artifact was produced by Positron AI from google/gemma-4-12B-it. The original model license and usage restrictions continue to apply.
