bekan/english_karakalpak_pairs_parallel_corpus_v2_8907
English-Karakalpak Parallel Corpus v2 (8.9K) Dataset Description English-Karakalpak Parallel Corpus v2 is a high-quality dataset containing 8,906 carefully aligned sentence pairs in English (en) and Karakalpak (kaa). This dataset is designed to advance the representation and capability of the Karakalpak language in large-scale AI models (LLMs) and Neural Machine Translation (NMT) systems, enabling them to better understand and generate Karakalpak text. This… See the full description on the dataset page: https://huggingface.co/datasets/bekan/english_karakalpak_pairs_parallel_corpus_v2_8907.
Update README.md
Upload en_kaa_parallel_corpus_v2_8907.csv
Delete en_kaa_v2_8907.csv
Update README.md
Update README.md
Edit from Data Studio
Edit from Data Studio
Update README.md
Update README.md
Initial commit
initial commit
