CoolFace
Modelpublic

LLMWildling/NVIDIA-Nemotron-3-Super-165B-A17B-Coder-NVFP4

sourceHugging Faceotherupdated 3mo agoView on Hugging Face
0likes68downloads
Model Card

NVIDIA-Nemotron-3-Super-165B-A17B-Coder-NVFP4

A higher-active-capacity coding variant of NVIDIA Nemotron 3 Super 120B-A12B NVFP4.

This is a community model and is not an official NVIDIA release. It was inspired by NVIDIA's open-model, open-data, and open-tooling work around Nemotron.

Model Summary

Total ParametersApproximately 165B
Active ParametersApproximately 17B per token
QuantizationNVFP4 mixed-precision checkpoint
ArchitectureNemotron hybrid Mamba-2, LatentMoE, Attention, and MTP
Validated Context Length131,072 tokens
SpecializationAgentic coding, reasoning, tool use, and multi-turn software-engineering workflows
Base ModelNVIDIA-Nemotron-3-Super-120B-A12B-NVFP4

Does This Work?

The public Nemotron 130B LLMWildling Canary NVFP4 provides a smaller proof point. It demonstrates direct recall of newly added domain knowledge and carries that knowledge into a follow-up task without RAG or prompt-injected context.

Intended Use

This checkpoint is intended for production coding assistants, repository analysis, agentic software-engineering systems, and tool-using workflows.

License

This model is derived from NVIDIA Nemotron 3 Super. Use is governed by the NVIDIA Nemotron Open Model License. Review the upstream model card for its full terms, safety information, limitations, and base-model details.

LLMWildling/NVIDIA-Nemotron-3-Super-165B-A17B-Coder-NVFP4 · CoolFace