forkjoin-ai/llama-3.1-nemotron-nano-vl-8b-v1-gguf
1126
Llama 3.1 Nemotron Nano Vl 8B V1
Forkjoin.ai conversion of nvidia/Llama-3.1-Nemotron-Nano-VL-8B-v1 to GGUF format for edge deployment.
Model Details
- Source Model: nvidia/Llama-3.1-Nemotron-Nano-VL-8B-v1
- Format: GGUF
- Converted by: Forkjoin.ai
Usage
With llama.cpp
./llama-cli -m nvidia_Llama-3.1-Nemotron-Nano-8B-v1-Q4_K_M.gguf -p "Your prompt here" -n 256With Ollama
Create a Modelfile:
FROM ./nvidia_Llama-3.1-Nemotron-Nano-8B-v1-Q4_K_M.ggufollama create llama-3.1-nemotron-nano-vl-8b-v1-gguf -f Modelfile
ollama run llama-3.1-nemotron-nano-vl-8b-v1-ggufAbout Forkjoin.ai
Forkjoin.ai runs AI models at the edge -- in-browser, on-device, zero cloud cost. These converted models power real-time inference, speech recognition, and natural language capabilities.
All conversions are optimized for edge deployment within browser and mobile memory constraints.
License
Apache 2.0 (follows upstream model license)
