shaowenchen/baichuan2-7b-base-gguf
3441
Provided files
Usage:
docker run --rm -it -p 8000:8000 -v /path/to/models:/models -e MODEL=/models/gguf-model-name.gguf hubimage/llama-cpp-python:latestand you can view http://localhost:8000/docs to see the swagger UI.
Provided images
Usage:
docker run --rm -p 8000:8000 shaowenchen/baichuan2-7b-base-gguf:Q2_Kand you can view http://localhost:8000/docs to see the swagger UI.
