CoolFace
Modelpublic

FunAudioLLM/Fun-ASR-Nano-GGUF

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
5likes6.3kdownloads
Model Card

Fun-ASR-Nano · GGUF (FunASR llama.cpp runtime)

GGUF build of Fun-ASR-Nano (SenseVoice SAN-M encoder + adaptor + Qwen3-0.6B LLM decoder) for the zero-Python, CPU/edge [FunASR llama.cpp runtime](https://github.com/FunAudioLLM/Fun-ASR/tree/main/runtime/llama.cpp) — the accuracy leader (LLM decoder), single C++ binary.

LLM quantization (pick by size vs accuracy)

The Fun-ASR-Nano LLM (Qwen3-0.6B) ships in three tiers — all within 0.1% CER (184-file micro-CER). Pair any with funasr-encoder-f16.gguf (470 MB).

LLM filesizeCER ↓speed
qwen3-0.6b-q4km.gguf484 MB8.35%6.1×smallest
qwen3-0.6b-q5km.gguf551 MB8.25%5.7×best accuracy
qwen3-0.6b-q8_0.gguf805 MB8.30%6.0×

Recommended: q4_K_M (smallest) or q5_K_M (best).

Get it running (no Python, no build)

These are GGUF weights for the [FunASR llama.cpp runtime](https://github.com/modelscope/FunASR/tree/main/runtime/llama.cpp) — a whisper.cpp-style, single self-contained binary for CPU / edge. Grab a prebuilt binary, then fetch this model and run:

  • Prebuilt binaries (Linux / macOS / Windows) → [GitHub Releases](https://github.com/modelscope/FunASR/releases) (tag runtime-llamacpp-v*)
  • Deployment guide & qualified benchmarks → [funasr.com/deploy/llama-cpp](https://www.funasr.com/deploy/llama-cpp.html)
bash
bash download-funasr-model.sh nano ./gguf
llama-funasr-cli --enc ./gguf/funasr-encoder-f16.gguf -m ./gguf/qwen3-0.6b-q8_0.gguf --vad ./gguf/fsmn-vad.gguf -a audio.wav

Files

filesizenotes
funasr-encoder-f16.gguf470 MBaudio encoder + adaptor (f16)
qwen3-0.6b-q8_0.gguf805 MBLLM decoder, recommended (Q8_0)
qwen3-0.6b-q4km.gguf484 MBLLM decoder, smaller (Q4KM)

Usage (needs both the encoder and the LLM gguf)

bash
llama-funasr-cli --enc funasr-encoder-f16.gguf -m qwen3-0.6b-q8_0.gguf -a audio.wav --vad fsmn-vad.gguf

On CPU: 8.30 % CER on the 184-clip Mandarin benchmark (vs whisper.cpp 22–31 %).

Links

  • 🧩 Runtime & build: [Fun-ASR · runtime/llama.cpp](https://github.com/FunAudioLLM/Fun-ASR/tree/main/runtime/llama.cpp) — ⭐ Star [Fun-ASR](https://github.com/FunAudioLLM/Fun-ASR)!
  • Source model: FunAudioLLM/Fun-ASR-Nano-2512