CodeStrux-Tech/tac-1-gguf
tac-1-gguf — GGUF format for llama.cpp and Ollama
Overview
tac-1-gguf provides GGUF files quantized from `CodeStrux-Tech/tac-1` using llama.cpp commit 67776ea.
Serving
Ollama
ollama run hf.co/CodeStrux-Tech/tac-1-gguf:Q5_K_MOllama registers an HF-pulled model under the full hf.co/...:Q5_K_M name. To use it with the tico client (default model name tac-1), alias it once with ollama cp hf.co/CodeStrux-Tech/tac-1-gguf:Q5_K_M tac-1, or set TICO_OLLAMA_MODEL="hf.co/CodeStrux-Tech/tac-1-gguf:Q5_K_M".
llama.cpp server
llama-server -m tac-1-Q5_K_M.gguf --jinjaThe server is OpenAI-compatible at /v1. Client env: OLLAMA_BASE_URL=http://localhost:8080/v1 TICO_OLLAMA_MODEL=tac-1.
Chat template warning
The chat template is ChatML with EOS id 151645 (<|im_end|>). There are 0 think references in the template. The template is byte-identical across the merged bf16, FP8, and GGUF builds. When using llama.cpp or Ollama, ensure the server applies the ChatML template correctly — the --jinja flag on llama-server enables this.
Training data attribution
Contains information from OpenStreetMap (https://www.openstreetmap.org/copyright), which is made available under the Open Database License (ODbL) 1.0. © OpenStreetMap contributors.
For full training details, architecture, evaluation, and limitations, see `CodeStrux-Tech/tac-1`.
tac-1 is a derivative work of Qwen/Qwen3-4B-Instruct-2507, Copyright 2024 Alibaba Cloud, licensed under the Apache License, Version 2.0. The upstream LICENSE is included in this repository.
