zhangsq-nju/EdgeRazor-PlayGround
<div align="center"> <br/> <img src="https://raw.githubusercontent.com/zhangsq-nju/EdgeRazor/main/asset/Logo-full.png" alt="EdgeRazor Logo" width="60%"> <h3> Lightweight Framework for Edge AI </h3>
<p> <a href="https://arxiv.org/abs/2605.04062" target="_blank"> <img src="https://img.shields.io/badge/arXiv-EdgeRazor-b31b1b?style=flat&logo=arxiv" alt="arXiv EdgeRazor"> </a> <a href="https://github.com/zhangsq-nju/EdgeRazor" target="blank"> <img src="https://img.shields.io/badge/GitHub-EdgeRazor-blue?style=flat&logo=github" alt="GitHub EdgeRazor"> </a> </p>
<h5> ✨ If you like our project, please consider giving us ❤️. </h5> </div>
EdgeRazor Playground
A CPU-friendly chatbot powered by [Qwen3-EdgeRazor-nbit](https://huggingface.co/collections/zhangsq-nju/edgerazor-nbit), running locally via llama.cpp. Displays real-time efficiency metrics (output tokens, time, decoding throughput) per turn.
Dependencies
- llama-cpp-python
- Qwen3-EdgeRazor-nbit gguf files:
- Qwen3-0.6B-EdgeRazor-GGUF
- Qwen3-1.7B-EdgeRazor-GGUF
