CoolFace
Modelpublic

aungzaythant/Raj-1.7B

sourceHugging Faceotherupdated 4d agoView on Hugging Face
0likes39downloads
Model Card

<img src="assets/raj-logo-card.png" width="120" height="120" alt="Raj logo">

Introducing Raj 1.7B

English · မြန်မာစာ

<a id="english"></a>

Burmese speech, in words. Now with room for English.

Meet Raj 1.7B, the larger model in the next generation of Raj. It turns spoken Burmese into text, including the English words, names, and phrases we mix into everyday conversations.

[image]

A good listener. Getting better.

From Nemotron to Qwen3-ASR

Raj-0 started with NVIDIA’s Nemotron ASR. For this generation, we’ve moved to Qwen3-ASR-1.7B as the base model, with a stronger focus on the way Burmese and English are spoken together.

What’s improved since Raj-0

  • —Burmese–English code-switching — better handling of English words and phrases within Burmese speech.
  • —Burmese transcription — more accurate recognition of everyday spoken Burmese.
  • —Spelling and names — more consistent spelling and better recognition of names and borrowed words.
  • —Spacing — cleaner, more readable transcripts.

A note on speed

These models can be a little slower than Raj-0, particularly Raj 1.7B. The tradeoff is improved transcription and code-switching. Speed depends on your device and the format you use.

What we’re working on next

For Raj 1.7B, our focus is transcription quality, especially for more challenging speech and recordings.

  • —Challenging speech — better recognition of fast speech, different accents, and recordings with background noise.
  • —Code-switching, names, and vocabulary — more accurate Burmese–English transitions, personal and place names, and specialized terms.
  • —Longer recordings — more consistent transcripts, with fewer missed or repeated passages.
  • —Faithful, readable text — better spelling, spacing, and punctuation while preserving what was actually said.

These are development priorities for future versions.

Downloads

Both formats are available in this repository for local inference.

FormatDownload
GGUF · 8-bitGGUF files
MLX · 8-bitMLX files

Raj 0.6B is the compact option. Raj 1.7B is the larger model.

Noncommercial use

For personal projects, learning, and noncommercial research. Commercial use is not allowed for now. See LICENSE and NOTICE. The original Qwen materials retain their upstream license.

Thanks for trying Raj. We’re still listening.


<a id="burmese"></a>

မြန်မာစာ

မြန်မာစကားကို စာသားအဖြစ်။ အင်္ဂလိပ်လို ရောပြောတာတွေလည်း ပါဝင်လာပါပြီ။

Raj ရဲ့ မျိုးဆက်သစ်ထဲက ပိုကြီးတဲ့ မော်ဒယ်ဖြစ်တဲ့ Raj 1.7B ကို မိတ်ဆက်ပေးချင်ပါတယ်။ မြန်မာလို ပြောထားတဲ့အသံတွေကို စာသားပြောင်းပေးပြီး နေ့စဉ်စကားပြောရာမှာ ညှပ်သုံးတဲ့ အင်္ဂလိပ်စကားလုံးတွေ၊ နာမည်တွေနဲ့ စကားစုတွေကိုလည်း ရေးထုတ်ပေးနိုင်ပါတယ်။

နားထောင်တတ်ဖို့ ဆက်လေ့လာနေဆဲပါ။

Nemotron ကနေ Qwen3-ASR သို့

ပထမဗားရှင်း Raj-0 ကို NVIDIA ရဲ့ Nemotron ASR ကို အခြေခံပြီး တည်ဆောက်ခဲ့ပါတယ်။ ဒီတစ်ခါမှာတော့ Qwen3-ASR-1.7B ကို ပြောင်းလဲအသုံးပြုထားပြီး မြန်မာနဲ့ အင်္ဂလိပ် ရောပြောတဲ့ စကားသံတွေကို ပိုကောင်းကောင်း နားလည်နိုင်ဖို့ ဦးစားပေးထားပါတယ်။

Raj-0 ထက် ဘာတွေ ပိုကောင်းလာလဲ

  • —မြန်မာ၊ အင်္ဂလိပ် ရောပြောခြင်း — မြန်မာစကားကြားမှာ ပါလာတဲ့ အင်္ဂလိပ်စကားလုံးနဲ့ စကားစုတွေကို ပိုကောင်းကောင်း ရေးထုတ်နိုင်လာပါတယ်။
  • —မြန်မာစကားကို စာသားပြောင်းခြင်း — နေ့စဉ်ပြောနေတဲ့ မြန်မာစကားတွေကို ပိုမှန်ကန်စွာ ရေးထုတ်နိုင်လာပါတယ်။
  • —စာလုံးပေါင်းနဲ့ နာမည်များ — စာလုံးပေါင်းတွေ ပိုမှန်လာပြီး နာမည်တွေနဲ့ မွေးစားစကားလုံးတွေကိုလည်း ပိုကောင်းကောင်း ခွဲခြားရေးထုတ်နိုင်လာပါတယ်။
  • —စာသားအကွာအဝေး — စကားလုံးနဲ့ စကားစုတွေကြား နေရာခြားပုံ ပိုသပ်ရပ်လာလို့ ဖတ်ရတာ ပိုလွယ်လာပါတယ်။

အမြန်နှုန်းအကြောင်း

ဒီမော်ဒယ်တွေက Raj-0 ထက် နည်းနည်းပိုနှေးနိုင်ပါတယ်။ အထူးသဖြင့် Raj 1.7B မှာ ပိုသိသာနိုင်ပါတယ်။ အဲဒီအစား မြန်မာစကားနဲ့ အင်္ဂလိပ်လို ရောပြောတာတွေကို ပိုမှန်ကန်စွာ ရေးထုတ်နိုင်ဖို့ ဦးစားပေးထားပါတယ်။ အမြန်နှုန်းက သုံးတဲ့စက်နဲ့ မော်ဒယ်ဖော်မတ်ပေါ်လည်း မူတည်ပါတယ်။

နောက်တစ်ဆင့် ဘာတွေ ဆက်လုပ်မလဲ

Raj 1.7B မှာတော့ စာသားတိကျမှန်ကန်မှုကို ဦးစားပေးသွားမှာပါ။ အထူးသဖြင့် နားထောင်ရခက်တဲ့ စကားနဲ့ အသံဖိုင်တွေကို ပိုကောင်းကောင်း ရေးထုတ်နိုင်ဖို့ ရည်ရွယ်ပါတယ်။

  • —နားထောင်ရခက်တဲ့ စကားတွေကို ပိုမှန်မှန် ရေးထုတ်နိုင်ဖို့ — ခပ်မြန်မြန် ပြောတာ၊ လေသံကွဲတာနဲ့ နောက်ခံဆူညံသံ ပါတာတွေကို ပိုကောင်းကောင်း ကိုင်တွယ်နိုင်ဖို့။
  • —ဘာသာစကား ရောပြောတာ၊ နာမည်နဲ့ အသုံးအနှုန်းတွေ ပိုမှန်ဖို့ — မြန်မာနဲ့ အင်္ဂလိပ် ရောပြောတာအပြင် လူနာမည်၊ နေရာနာမည်နဲ့ နယ်ပယ်အလိုက် သုံးတဲ့ စကားလုံးတွေကို ပိုတိတိကျကျ ရေးထုတ်နိုင်ဖို့။
  • —အသံဖိုင်ရှည်တွေမှာ ပိုတည်ငြိမ်ဖို့ — စကားတွေ ကျန်ခဲ့တာ၊ ထပ်ရေးမိတာ နည်းလာပြီး အစအဆုံး ပိုမှန်မှန် ရေးထုတ်နိုင်ဖို့။
  • —ပြောတဲ့အတိုင်း မှန်ပြီး ဖတ်ရလွယ်ဖို့ — မူရင်းပြောခဲ့တဲ့ စကားကို မပြောင်းဘဲ စာလုံးပေါင်း၊ စာလုံးအကွာအဝေးနဲ့ ပုဒ်ဖြတ်ပုဒ်ရပ်တွေ ပိုကောင်းလာဖို့။

ဒါတွေက နောက်ထွက်မယ့် ဗားရှင်းတွေအတွက် ဦးစားပေးလုပ်သွားမယ့် အရာတွေပါ။

မော်ဒယ်ဖိုင်များ

ကိုယ့်စက်ပေါ်မှာ အသုံးပြုနိုင်ဖို့ ဖော်မတ်နှစ်မျိုးစလုံးကို ဒီ repository မှာ ရယူနိုင်ပါတယ်။

ဖော်မတ်မော်ဒယ်ဖိုင်များ
GGUF · 8-bitGGUF ဖိုင်များ
MLX · 8-bitMLX ဖိုင်များ

Raj 0.6B က ပိုပေါ့ပါးတဲ့ မော်ဒယ်ဖြစ်ပြီး Raj 1.7B က ပိုကြီးတဲ့ မော်ဒယ်ပါ။

အသုံးပြုခွင့်

ကိုယ်ပိုင်ပရောဂျက်တွေ၊ လေ့လာသင်ယူမှုနဲ့ စီးပွားရေးရည်ရွယ်ချက်မပါတဲ့ သုတေသနတွေအတွက် သုံးနိုင်ပါတယ်။ စီးပွားရေးအတွက် အသုံးပြုတာကိုတော့ လက်ရှိမှာ ခွင့်မပြုသေးပါဘူး။ အသေးစိတ်ကို LICENSE နဲ့ NOTICE မှာ ဖတ်နိုင်ပါတယ်။ မူရင်း Qwen ဖိုင်တွေအတွက် မူရင်းလိုင်စင် က ဆက်လက်အကျုံးဝင်ပါတယ်။

Raj ကို စမ်းသုံးပေးတဲ့အတွက် ကျေးဇူးတင်ပါတယ်။ ပိုကောင်းလာအောင် ဆက်နားထောင်သွားပါမယ်။