CoolFace
Datasetpublic

failed09/bashkir-frequency-index

Bashkir Frequency Index v11.5 Word-frequency index for Bashkir, computed over a large monolingual Bashkir-language dataset, for NLP, spellchecking and lexical research. Overview Word-frequency index for the Bashkir language computed over a large monolingual Bashkir-language dataset. Non-Bashkir admixture, borrowed vocabulary and scanning artifacts were reduced with automated language filtering. The public configuration (count ≥ 3) is the recommended default;… See the full description on the dataset page: https://huggingface.co/datasets/failed09/bashkir-frequency-index.

sourceHugging Facecc-by-4.0updated 3d agoView on Hugging Face
0likes238downloads
settings

This repository belongs to failed09 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namebashkir-frequency-index
visibilitypublic
licencecc-by-4.0
gatedno
ownerfailed09
Account settings
failed09/bashkir-frequency-index · CoolFace