datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
rundi-tumbuka_sentence-pairs
Rundi-Tumbuka_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Tumbuka_Sentence-Pairs
Number of Rows: 194527
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-tumbuka_sentence-pairs.ewe-rundi_sentence-pairs
Ewe-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Ewe-Rundi_Sentence-Pairs
Number of Rows: 197511
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/ewe-rundi_sentence-pairs.fulah-rundi_sentence-pairs
Fulah-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Rundi_Sentence-Pairs
Number of Rows: 80036
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-rundi_sentence-pairs.hausa-rundi_sentence-pairs
Hausa-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Hausa-Rundi_Sentence-Pairs
Number of Rows: 261557
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/hausa-rundi_sentence-pairs.rundi-swahili_sentence-pairs
Rundi-Swahili_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Swahili_Sentence-Pairs
Number of Rows: 622955
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-swahili_sentence-pairs.kimbundu-rundi_sentence-pairs
Kimbundu-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Kimbundu-Rundi_Sentence-Pairs
Number of Rows: 75678
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kimbundu-rundi_sentence-pairs.rundi-somali_sentence-pairs
Rundi-Somali_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Somali_Sentence-Pairs
Number of Rows: 210469
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-somali_sentence-pairs.rundi-twi_sentence-pairs
Rundi-Twi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Twi_Sentence-Pairs
Number of Rows: 199589
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-twi_sentence-pairs.oromo-rundi_sentence-pairs
Oromo-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Oromo-Rundi_Sentence-Pairs
Number of Rows: 91889
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/oromo-rundi_sentence-pairs.akan-rundi_sentence-pairs
Akan-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Rundi_Sentence-Pairs
Number of Rows: 47706
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-rundi_sentence-pairs.rundi-xhosa_sentence-pairs
Rundi-Xhosa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Xhosa_Sentence-Pairs
Number of Rows: 307505
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-xhosa_sentence-pairs.rundi-yoruba_sentence-pairs
Rundi-Yoruba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Yoruba_Sentence-Pairs
Number of Rows: 174926
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-yoruba_sentence-pairs.kongo-rundi_sentence-pairs
Kongo-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Kongo-Rundi_Sentence-Pairs
Number of Rows: 99480
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kongo-rundi_sentence-pairs.dyula-rundi_sentence-pairs
Dyula-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Dyula-Rundi_Sentence-Pairs
Number of Rows: 72663
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/dyula-rundi_sentence-pairs.bemba-rundi_sentence-pairs
Bemba-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bemba-Rundi_Sentence-Pairs
Number of Rows: 208683
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bemba-rundi_sentence-pairs.rundi-tsonga_sentence-pairs
Rundi-Tsonga_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Tsonga_Sentence-Pairs
Number of Rows: 283463
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-tsonga_sentence-pairs.kikuyu-rundi_sentence-pairs
Kikuyu-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Kikuyu-Rundi_Sentence-Pairs
Number of Rows: 59660
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kikuyu-rundi_sentence-pairs.fon-rundi_sentence-pairs
Fon-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fon-Rundi_Sentence-Pairs
Number of Rows: 89002
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fon-rundi_sentence-pairs.dinka-rundi_sentence-pairs
Dinka-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Dinka-Rundi_Sentence-Pairs
Number of Rows: 24413
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/dinka-rundi_sentence-pairs.rundi-shona_sentence-pairs
Rundi-Shona_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Shona_Sentence-Pairs
Number of Rows: 334206
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-shona_sentence-pairs.rundi-wolof_sentence-pairs
Rundi-Wolof_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Wolof_Sentence-Pairs
Number of Rows: 45932
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-wolof_sentence-pairs.rundi-tigrinya_sentence-pairs
Rundi-Tigrinya_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Tigrinya_Sentence-Pairs
Number of Rows: 161796
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-tigrinya_sentence-pairs.chichewa-rundi_sentence-pairs
Chichewa-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Chichewa-Rundi_Sentence-Pairs
Number of Rows: 339947
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/chichewa-rundi_sentence-pairs.rundi-umbundu_sentence-pairs
Rundi-Umbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Umbundu_Sentence-Pairs
Number of Rows: 121294
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-umbundu_sentence-pairs.igbo-rundi_sentence-pairs
Igbo-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Igbo-Rundi_Sentence-Pairs
Number of Rows: 134613
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/igbo-rundi_sentence-pairs.rundi-zulu_sentence-pairs
Rundi-Zulu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Zulu_Sentence-Pairs
Number of Rows: 467223
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-zulu_sentence-pairs.rundi-tswana_sentence-pairs
Rundi-Tswana_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Tswana_Sentence-Pairs
Number of Rows: 221203
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-tswana_sentence-pairs.rundi-swati_sentence-pairs
Rundi-Swati_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Rundi-Swati_Sentence-Pairs
Number of Rows: 71808
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-swati_sentence-pairs.pedi-rundi_sentence-pairs
Pedi-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Pedi-Rundi_Sentence-Pairs
Number of Rows: 142452
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/pedi-rundi_sentence-pairs.lingala-rundi_sentence-pairs
Lingala-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Lingala-Rundi_Sentence-Pairs
Number of Rows: 180158
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/lingala-rundi_sentence-pairs.
