CoolFace
Datasetpublic

BeitTigreAI/tigre-data-kenLM

Tigre 5-gram Language Model (KenLM) Overview This repository provides a 5-gram Language Model (LM) for the Tigre language, trained using the KenLM toolkit. This model is a foundational resource for various downstream NLP and speech applications, including: Rescoring hypotheses in Automatic Speech Recognition (ASR). Improving text generation and fluency in Machine Translation (MT). Performing basic text filtering and quality control. The model is provided in the… See the full description on the dataset page: https://huggingface.co/datasets/BeitTigreAI/tigre-data-kenLM.

sourceHugging Facecc-by-sa-4.0updated 19d agoView on Hugging Face
0likes47downloads
tigre-data-kenLM.binary4 linesDownload Raw Back to root
1version https://git-lfs.github.com/spec/v12oid sha256:1cd2dbb9c97188018a96e208930872c5bec1a5abcfc9ddf246d6be6aba9d95853size 5939911364