CoolFace
Modelpublic

JoaoZaokk/kotoba-whisper-v2.0-ggml

sourceHugging Faceapache-2.0updated 15d agoView on Hugging Face
1likes
Model Card

kotoba-whisper-v2.0-ggml

GGML conversion of kotoba-tech/kotoba-whisper-v2.0 (a Whisper checkpoint) for whisper.cpp, in f16 plus q40 / q50 / q8_0 quantizations.

Source checkpoint: kotoba-tech/kotoba-whisper-v2.0 by kotoba-tech · License: apache-2.0 (unchanged; this repo only re-packages the weights) Engine: Load with whisper.cpp (whisper-cli -m <file>), or any app that embeds it.

Files

FileQuantizationSizeNote
ggml-kotoba-whisper-v2.0-f16.binf161520 MB
ggml-kotoba-whisper-v2.0-q4_0.binq4_0444 MB
ggml-kotoba-whisper-v2.0-q5_0.binq5_0538 MB
ggml-kotoba-whisper-v2.0-q8_0.binq8_0818 MB

f16 is the lossless conversion; q8_0 is nearly identical in accuracy at ~55 % of the size; q5_0/q5_k are the phone-friendly choice; q4_* is smallest with a small accuracy cost.

How these were made

Converted from the upstream checkpoint with the engine's own converter, then quantized with the engine's quantizer. Each variant was checked by transcribing short Portuguese and English samples before upload.

Attribution

Weights are derivative works of the upstream model and keep its license. Please cite the original authors (kotoba-tech). The conversion and hosting here are maintained by JoaoZaokk so that the download links used by the Odysseus / Open WebUI native apps stay stable. No warranty.