shaowenchen/longchat-13b-16k-gguf
1433
Provided files
Usage:
docker run --rm -it -p 8000:8000 -v /path/to/models:/models -e MODEL=/models/gguf-model-name.gguf hubimage/llama-cpp-python:latestand you can view http://localhost:8000/docs to see the swagger UI.
Provided images
Usage:
docker run --rm -p 8000:8000 shaowenchen/longchat-13b-16k-gguf:Q2_Kand you can view http://localhost:8000/docs to see the swagger UI.
