forkjoin-ai/nvidia-nemotron-3-nano-30b-a3b-bf16-gguf
020
Nvidia Nemotron 3 Nano 30B A3B Bf16
Forkjoin.ai conversion of nvidia/Nemotron-3-Nano-30B-A3B-BF16 to GGUF format for edge deployment.
Model Details
- Source Model: nvidia/Nemotron-3-Nano-30B-A3B-BF16
- Format: GGUF
- Converted by: Forkjoin.ai
Usage
With llama.cpp
./llama-cli -m nvidia_Nemotron-3-Nano-30B-A3B-Q4_K_M.gguf -p "Your prompt here" -n 256With Ollama
Create a Modelfile:
FROM ./nvidia_Nemotron-3-Nano-30B-A3B-Q4_K_M.ggufollama create nvidia-nemotron-3-nano-30b-a3b-bf16-gguf -f Modelfile
ollama run nvidia-nemotron-3-nano-30b-a3b-bf16-ggufAbout Forkjoin.ai
Forkjoin.ai runs AI models at the edge -- in-browser, on-device, zero cloud cost. These converted models power real-time inference, speech recognition, and natural language capabilities.
All conversions are optimized for edge deployment within browser and mobile memory constraints.
License
Apache 2.0 (follows upstream model license)
