CoolFace
Modelpublic

songlindotiot/chunkformer-large-vie

sourceHugging Facecc-by-nc-4.0updated 1y agoView on Hugging Face
0likes2downloads
Model Card

ChunkFormer-Large-Vie: Large-Scale Pretrained ChunkFormer for Vietnamese Automatic Speech Recognition

<style> img { display: inline; } </style> ![Ranked #1: Speech Recognition on Common Voice Vi](https://paperswithcode.com/sota/speech-recognition-on-common-voice-vi) ![Ranked #1: Speech Recognition on VIVOS](https://paperswithcode.com/sota/speech-recognition-on-vivos)

![License: CC BY-NC 4.0](https://creativecommons.org/licenses/by-nc/4.0/) ![GitHub](https://github.com/khanld/chunkformer) ![Paper](https://arxiv.org/abs/2502.14673) ![Model size](#description)


<a name = "citation" ></a>

Citation

If you use this work in your research, please cite:

bibtex
@INPROCEEDINGS{10888640,
  author={Le, Khanh and Ho, Tuan Vu and Tran, Dung and Chau, Duc Thanh},
  booktitle={ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)}, 
  title={ChunkFormer: Masked Chunking Conformer For Long-Form Speech Transcription}, 
  year={2025},
  volume={},
  number={},
  pages={1-5},
  keywords={Scalability;Memory management;Graphics processing units;Signal processing;Performance gain;Hardware;Resource management;Speech processing;Standards;Context modeling;chunkformer;masked batch;long-form transcription},
  doi={10.1109/ICASSP49660.2025.10888640}}
}

<a name = "contact"></a>

Contact

  • —khanhld218@gmail.com
  • —![GitHub](https://github.com/khanld)
  • —![LinkedIn](https://www.linkedin.com/in/khanhld257/)