ggml-org/DeepSeek-V4-Flash-GGUF
111.4k
DeepSeek-V4-Flash
Run with https://llama.app
llama serve -hf ggml-org/DeepSeek-V4-Flash-GGUFSource models
- https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash
Notes
- Currently, the Q2 models do not use an imatrix calibration due to lack of one.
TODOs
- add info
[!IMPORTANT] This model is automatically converted using https://github.com/ggml-org/convert
