datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bambara-fon_sentence-pairs
Bambara-Fon_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Fon_Sentence-Pairs
Number of Rows: 25525
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-fon_sentence-pairs.bambara-somali_sentence-pairs
Bambara-Somali_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Somali_Sentence-Pairs
Number of Rows: 71330
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-somali_sentence-pairs.bambara-kongo_sentence-pairs
Bambara-Kongo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Kongo_Sentence-Pairs
Number of Rows: 24663
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-kongo_sentence-pairs.bambara-xhosa_sentence-pairs
Bambara-Xhosa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Xhosa_Sentence-Pairs
Number of Rows: 47407
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-xhosa_sentence-pairs.bambara-tigrinya_sentence-pairs
Bambara-Tigrinya_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Tigrinya_Sentence-Pairs
Number of Rows: 28792… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-tigrinya_sentence-pairs.bambara-swati_sentence-pairs
Bambara-Swati_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Swati_Sentence-Pairs
Number of Rows: 15607
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-swati_sentence-pairs.bambara-lingala_sentence-pairs
Bambara-Lingala_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Lingala_Sentence-Pairs
Number of Rows: 32701
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-lingala_sentence-pairs.bambara-igbo_sentence-pairs
Bambara-Igbo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Igbo_Sentence-Pairs
Number of Rows: 37303
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-igbo_sentence-pairs.afrikaans-bambara_sentence-pairs
Afrikaans-Bambara_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Afrikaans-Bambara_Sentence-Pairs
Number of Rows: 121709… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/afrikaans-bambara_sentence-pairs.bambara-shona_sentence-pairs
Bambara-Shona_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Shona_Sentence-Pairs
Number of Rows: 57467
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-shona_sentence-pairs.bambara-kikuyu_sentence-pairs
Bambara-Kikuyu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Kikuyu_Sentence-Pairs
Number of Rows: 16933
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-kikuyu_sentence-pairs.bambara-dinka_sentence-pairs
Bambara-Dinka_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Dinka_Sentence-Pairs
Number of Rows: 13958
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-dinka_sentence-pairs.bambara-yoruba_sentence-pairs
Bambara-Yoruba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Yoruba_Sentence-Pairs
Number of Rows: 61950
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-yoruba_sentence-pairs.bambara-twi_sentence-pairs
Bambara-Twi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Twi_Sentence-Pairs
Number of Rows: 44143
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-twi_sentence-pairs.bambara-kimbundu_sentence-pairs
Bambara-Kimbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Kimbundu_Sentence-Pairs
Number of Rows: 13259… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-kimbundu_sentence-pairs.bambara-swahili_sentence-pairs
Bambara-Swahili_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Swahili_Sentence-Pairs
Number of Rows: 146220
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-swahili_sentence-pairs.bambara-kamba_sentence-pairs
Bambara-Kamba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Kamba_Sentence-Pairs
Number of Rows: 12455
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-kamba_sentence-pairs.bambara-dyula_sentence-pairs
Bambara-Dyula_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Dyula_Sentence-Pairs
Number of Rows: 23484
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-dyula_sentence-pairs.Bambara_sentimentbambara-wolof_sentence-pairs
Bambara-Wolof_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Wolof_Sentence-Pairs
Number of Rows: 20330
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-wolof_sentence-pairs.bambara-tsonga_sentence-pairs
Bambara-Tsonga_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Tsonga_Sentence-Pairs
Number of Rows: 41626
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-tsonga_sentence-pairs.bambara-hausa_sentence-pairs
Bambara-Hausa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Hausa_Sentence-Pairs
Number of Rows: 76988
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-hausa_sentence-pairs.bambara-bemba_sentence-pairs
Bambara-Bemba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Bemba_Sentence-Pairs
Number of Rows: 30900
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-bemba_sentence-pairs.bambara-umbundu_sentence-pairs
Bambara-Umbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Umbundu_Sentence-Pairs
Number of Rows: 20038
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-umbundu_sentence-pairs.bambara-tswana_sentence-pairs
Bambara-Tswana_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Tswana_Sentence-Pairs
Number of Rows: 55246
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-tswana_sentence-pairs.bambara-oromo_sentence-pairs
Bambara-Oromo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Oromo_Sentence-Pairs
Number of Rows: 30267
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-oromo_sentence-pairs.bambara-nuer_sentence-pairs
Bambara-Nuer_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Nuer_Sentence-Pairs
Number of Rows: 12783
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-nuer_sentence-pairs.bambara-chichewa_sentence-pairs
Bambara-Chichewa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Chichewa_Sentence-Pairs
Number of Rows: 52884… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-chichewa_sentence-pairs.amharic-bambara_sentence-pairs
Amharic-Bambara_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Amharic-Bambara_Sentence-Pairs
Number of Rows: 51636
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/amharic-bambara_sentence-pairs.bambara-zulu_sentence-pairs
Bambara-Zulu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bambara-Zulu_Sentence-Pairs
Number of Rows: 66614
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-zulu_sentence-pairs.
