CoolFace
Modelpublic

KasuleTrevor/cdli-qwen3-asr-lg-typical-textpre-0p6b-v2-finetune

sourceHugging Faceapache-2.0updated 4mo agoView on Hugging Face
0likes58downloads
Model Card

CDLI Qwen3-ASR Luganda Typical Speech Fine-tune

This repo contains the selected checkpoint checkpoint-20500 from the LG-QWEN3-ASR-TYPICAL-TEXTPRE-0P6B-V2-STAGE1 run.

Training Setup

  • —Base model: KasuleTrevor/cdli-qwen3-asr-lg-textpre-0p6b-v2
  • —Datasets: KasuleTrevor/lg_100hrs + dmusingu/yogera-dataset (Luganda split)
  • —Training language tag: Luganda
  • —Forced inference language: disabled
  • —Epochs: 5
  • —Batch size: 8
  • —Gradient accumulation: 2
  • —Learning rate: 2e-05
  • —Scheduler: cosine
  • —Warmup ratio: 0.03
  • —Save steps: 500
  • —Selected checkpoint: checkpoint-20500
  • —Selection reason: best validation normalized WER (0.218026, CER 0.048474)

Final Test Metrics

  • —Corpus WER (normalized): 0.265987
  • —Corpus CER (normalized): 0.069419
  • —Average utterance WER (normalized): 0.270043
  • —Average utterance CER (normalized): 0.070111

Checkpoint Selection Evidence

checkpointwer_normalizedcer_normalizedavg_wer_normalizedavg_cer_normalizedeval_loss
checkpoint-205000.21802563763710850.048473911826016620.244854611100630120.051529574509966470.3063507378101349
checkpoint-210000.21810492929826880.0484311315764910860.245452199597644030.0514823374307409250.30632004141807556
checkpoint-200000.21821065151314920.048726704209576550.244459393429687980.0515142170551013340.30641767382621765
checkpoint-190000.218422095942910.048540026757101520.24489931327129720.0514027628337014740.3062317371368408
checkpoint-210950.21847495705035020.0488472667309666760.245673166230260030.051655228144650060.30632004141807556
checkpoint-195000.219479318091714030.0490806135465604660.245685510361291770.051671330296247530.306130051612854

Source Dataset Breakdown

source_datasetn_samplesmean_wermean_cermedian_wermedian_cer
lg10028280.270.07010.22220.0379

Artifacts

  • —Result folder: results/checkpoint-20500/
  • —Includes checkpoint validation summaries, final test predictions, scored outputs, and grouped analyses.