datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
akan-umbundu_sentence-pairs
Akan-Umbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Umbundu_Sentence-Pairs
Number of Rows: 22651
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-umbundu_sentence-pairs.akan-oromo_sentence-pairs
Akan-Oromo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Oromo_Sentence-Pairs
Number of Rows: 47083
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-oromo_sentence-pairs.akan-tswana_sentence-pairs
Akan-Tswana_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Tswana_Sentence-Pairs
Number of Rows: 54079
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-tswana_sentence-pairs.akan-tsonga_sentence-pairs
Akan-Tsonga_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Tsonga_Sentence-Pairs
Number of Rows: 51276
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-tsonga_sentence-pairs.akan-swati_sentence-pairs
Akan-Swati_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Swati_Sentence-Pairs
Number of Rows: 19251
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-swati_sentence-pairs.akan-kinyarwanda_sentence-pairs
Akan-Kinyarwanda_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Kinyarwanda_Sentence-Pairs
Number of Rows: 67145… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-kinyarwanda_sentence-pairs.akan-somali_sentence-pairs
Akan-Somali_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Somali_Sentence-Pairs
Number of Rows: 78264
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-somali_sentence-pairs.akan-rundi_sentence-pairs
Akan-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Rundi_Sentence-Pairs
Number of Rows: 47706
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-rundi_sentence-pairs.akan-ewe_sentence-pairs
Akan-Ewe_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Ewe_Sentence-Pairs
Number of Rows: 64271
Number of Columns: 3… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-ewe_sentence-pairs.akan-twi_sentence-pairs
Akan-Twi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Twi_Sentence-Pairs
Number of Rows: 76666
Number of Columns: 3… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-twi_sentence-pairs.akan-dinka_sentence-pairs
Akan-Dinka_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Dinka_Sentence-Pairs
Number of Rows: 13582
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-dinka_sentence-pairs.akan-bemba_sentence-pairs
Akan-Bemba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Bemba_Sentence-Pairs
Number of Rows: 34555
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-bemba_sentence-pairs.akan-swahili_sentence-pairs
Akan-Swahili_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Swahili_Sentence-Pairs
Number of Rows: 165355
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-swahili_sentence-pairs.akan-pedi_sentence-pairs
Akan-Pedi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Pedi_Sentence-Pairs
Number of Rows: 46138
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-pedi_sentence-pairs.afrikaans-akan_sentence-pairs
Afrikaans-Akan_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Afrikaans-Akan_Sentence-Pairs
Number of Rows: 96786
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/afrikaans-akan_sentence-pairs.akan-kikuyu_sentence-pairs
Akan-Kikuyu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Kikuyu_Sentence-Pairs
Number of Rows: 13930
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-kikuyu_sentence-pairs.akan-kamba_sentence-pairs
Akan-Kamba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Kamba_Sentence-Pairs
Number of Rows: 11522
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-kamba_sentence-pairs.akan-igbo_sentence-pairs
Akan-Igbo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Igbo_Sentence-Pairs
Number of Rows: 39249
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-igbo_sentence-pairs.akan-hausa_sentence-pairs
Akan-Hausa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Hausa_Sentence-Pairs
Number of Rows: 98881
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-hausa_sentence-pairs.akan-zulu_sentence-pairs
Akan-Zulu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Zulu_Sentence-Pairs
Number of Rows: 82123
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-zulu_sentence-pairs.akan-kimbundu_sentence-pairs
Akan-Kimbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Kimbundu_Sentence-Pairs
Number of Rows: 15566
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-kimbundu_sentence-pairs.akan-ganda_sentence-pairs
Akan-Ganda_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Ganda_Sentence-Pairs
Number of Rows: 45950
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-ganda_sentence-pairs.akan-chichewa_sentence-pairs
Akan-Chichewa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Chichewa_Sentence-Pairs
Number of Rows: 52365
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-chichewa_sentence-pairs.akan-amharic_sentence-pairs
Akan-Amharic_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Amharic_Sentence-Pairs
Number of Rows: 101653
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-amharic_sentence-pairs.akan-xhosa_sentence-pairs
Akan-Xhosa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Xhosa_Sentence-Pairs
Number of Rows: 70567
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-xhosa_sentence-pairs.akan-wolof_sentence-pairs
Akan-Wolof_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Wolof_Sentence-Pairs
Number of Rows: 39026
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-wolof_sentence-pairs.akan-shona_sentence-pairs
Akan-Shona_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Shona_Sentence-Pairs
Number of Rows: 68092
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-shona_sentence-pairs.akan-nuer_sentence-pairs
Akan-Nuer_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Nuer_Sentence-Pairs
Number of Rows: 8534
Number of Columns: 3… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-nuer_sentence-pairs.akan-fulah_sentence-pairs
Akan-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Fulah_Sentence-Pairs
Number of Rows: 14234
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-fulah_sentence-pairs.akan-kongo_sentence-pairs
Akan-Kongo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Kongo_Sentence-Pairs
Number of Rows: 23039
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-kongo_sentence-pairs.
