CoolFace
Datasetpublic

BeitTigreAI/tigre-data-kenLM

Tigre 5-gram Language Model (KenLM) Overview This repository provides a 5-gram Language Model (LM) for the Tigre language, trained using the KenLM toolkit. This model is a foundational resource for various downstream NLP and speech applications, including: Rescoring hypotheses in Automatic Speech Recognition (ASR). Improving text generation and fluency in Machine Translation (MT). Performing basic text filtering and quality control. The model is provided in the… See the full description on the dataset page: https://huggingface.co/datasets/BeitTigreAI/tigre-data-kenLM.

sourceHugging Facecc-by-sa-4.0updated 17d agoView on Hugging Face
0likes45downloads

BeitTigreAI/tigre-data-kenLM · main · files are served by the source, never re-hosted here