junmingg/qwen2.5-coder-7b-text2sql-GGUF
0137
Qwen2.5-Coder-7B Text-to-SQL — GGUF quants
GGUF quantizations of `junmingg/qwen2.5-coder-7b-text2sql` for CPU/GPU inference via llama.cpp, Ollama, LM Studio, etc.
See the main model card for results (exact 78.8% / semantic 86.2% / validity 99.2% vs base 3.8 / 67.0 / 100), training details, and the required system prompt — the model expects the text-to-SQL system message + ChatML format.
Quick start (Ollama)
ollama run hf.co/junmingg/qwen2.5-coder-7b-text2sql-GGUF:Q4_K_MQuick start (llama.cpp)
llama-cli -hf junmingg/qwen2.5-coder-7b-text2sql-GGUF:Q4_K_MLicense: Apache-2.0 (base model) / data CC-BY-4.0.
