FermionResearch/Phonon-1-Big
6333
Phonon-1 Big
This is the largest build of the Phonon-1 family, an open speech recognition model for English that downloads in 581 MB.
Benchmarks
Word error rate, lower is better. Measured by us — full test sets, Whisper English text normalizer, greedy decoding.
Run it
pip install fermion-research
fermion transcribe recording.wav --model FermionResearch/Phonon-1-BigOr serve an OpenAI-compatible endpoint:
fermion serve --model FermionResearch/Phonon-1-Big
curl -s http://127.0.0.1:8000/v1/audio/transcriptions \
-F "file=@recording.wav" \
-F "model=FermionResearch/Phonon-1-Big"The same weights run on a Mac (via MLX), on an NVIDIA GPU, or on a plain CPU; the runtimes and Docker images are in the GitHub repo.
License
Apache License 2.0 for the weights and the command line. Base model: `Qwen/Qwen3-ASR-0.6B`, Apache-2.0.
