Phaedrus33/nemotron-speech-streaming-children-v17
033
nemotron-speech-streaming-en-0.6b - Fine-tuned on Children's Speech
Fine-tuned from nvidia/nemotron-speech-streaming-en-0.6b on children's speech data from the DrivenData ASR competition.
Training config
- Mode: full
- Epochs: 5
- Batch size: 16 (accumulate: 2)
- Learning rate: 1e-05
- Precision: bf16-mixed
- Speed perturbation: True
Usage
import nemo.collections.asr as nemo_asr
model = nemo_asr.models.ASRModel.restore_from("best_model.nemo")
hypotheses = model.transcribe(["audio.flac"])
print(hypotheses[0].text)