dasvad/open3dvqa-qwen3vl-4b-distill-q4-k-m-gguf
199
Open3DVQA Qwen3-VL 4B Distilled Q4KM GGUF
This repository contains the deployment files for the distilled Open3DVQA Qwen3-VL 4B student model.
Files
student_4b_merged-Q4_K_M.gguf Q4_K_M language model, about 2.4 GB
mmproj-student_4b_merged-f16.gguf F16 vision encoder/projector, about 798 MB
Modelfile.ollama Ollama import configuration
CODEX_ORIN_DEPLOY_GUIDE.md Detailed Jetson Orin NX instructionsBoth GGUF files are required for image inference.
Ollama
ollama create open3dvqa-qwen3vl:4b-q4km -f Modelfile.ollamaUse Ollama's /api/chat endpoint with base64 image data in messages[].images.
See CODEX_ORIN_DEPLOY_GUIDE.md for JetPack 5 deployment, checksums, GPU verification, API examples, and troubleshooting.
