CoolFace
Datasetpublicgated

BashkirNLPWorld/bashkir-russian-parallel

Dataset Card for Bashkir-Russian Parallel Corpus Dataset Details Dataset Description Bashkir-Russian Parallel Corpus is a large-scale sentence-aligned parallel corpus for the Bashkir–Russian language pair, assembled from authentic human-created translations. It contains 3,040,085 unique parallel sentence pairs, where each Bashkir sentence is aligned with its Russian counterpart. The corpus combines data from three open parallel corpora: TIL-MT… See the full description on the dataset page: https://huggingface.co/datasets/BashkirNLPWorld/bashkir-russian-parallel.

sourceHugging Faceotherupdated 5d agoView on Hugging Face
0likes27downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.