datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fulah-hausa_sentence-pairs
Fulah-Hausa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Hausa_Sentence-Pairs
Number of Rows: 269337
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-hausa_sentence-pairs.fulah-twi_sentence-pairs
Fulah-Twi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Twi_Sentence-Pairs
Number of Rows: 77858
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-twi_sentence-pairs.fulah-rundi_sentence-pairs
Fulah-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Rundi_Sentence-Pairs
Number of Rows: 80036
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-rundi_sentence-pairs.amharic-fulah_sentence-pairs
Amharic-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Amharic-Fulah_Sentence-Pairs
Number of Rows: 435048
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/amharic-fulah_sentence-pairs.fulah-wolof_sentence-pairs
Fulah-Wolof_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Wolof_Sentence-Pairs
Number of Rows: 46967
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-wolof_sentence-pairs.fulah-nuer_sentence-pairs
Fulah-Nuer_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Nuer_Sentence-Pairs
Number of Rows: 25702
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-nuer_sentence-pairs.fulah-chichewa_sentence-pairs
Fulah-Chichewa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Chichewa_Sentence-Pairs
Number of Rows: 220802
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-chichewa_sentence-pairs.fulah-swahili_sentence-pairs
Fulah-Swahili_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Swahili_Sentence-Pairs
Number of Rows: 323996
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-swahili_sentence-pairs.fulah-shona_sentence-pairs
Fulah-Shona_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Shona_Sentence-Pairs
Number of Rows: 122877
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-shona_sentence-pairs.fulah-oromo_sentence-pairs
Fulah-Oromo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Oromo_Sentence-Pairs
Number of Rows: 92794
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-oromo_sentence-pairs.fulah-ganda_sentence-pairs
Fulah-Ganda_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Ganda_Sentence-Pairs
Number of Rows: 91895
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-ganda_sentence-pairs.ewe-fulah_sentence-pairs
Ewe-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Ewe-Fulah_Sentence-Pairs
Number of Rows: 78099
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/ewe-fulah_sentence-pairs.afrikaans-fulah_sentence-pairs
Afrikaans-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Afrikaans-Fulah_Sentence-Pairs
Number of Rows: 168995
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/afrikaans-fulah_sentence-pairs.fulah-kamba_sentence-pairs
Fulah-Kamba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Kamba_Sentence-Pairs
Number of Rows: 14631
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-kamba_sentence-pairs.bemba-fulah_sentence-pairs
Bemba-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Bemba-Fulah_Sentence-Pairs
Number of Rows: 48062
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bemba-fulah_sentence-pairs.fulah-yoruba_sentence-pairs
Fulah-Yoruba_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Yoruba_Sentence-Pairs
Number of Rows: 374873
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-yoruba_sentence-pairs.fulah-tswana_sentence-pairs
Fulah-Tswana_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Tswana_Sentence-Pairs
Number of Rows: 88947
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-tswana_sentence-pairs.fulah-tigrinya_sentence-pairs
Fulah-Tigrinya_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Tigrinya_Sentence-Pairs
Number of Rows: 124029
Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-tigrinya_sentence-pairs.fulah-pedi_sentence-pairs
Fulah-Pedi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Pedi_Sentence-Pairs
Number of Rows: 55965
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-pedi_sentence-pairs.fulah-kimbundu_sentence-pairs
Fulah-Kimbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Kimbundu_Sentence-Pairs
Number of Rows: 19095
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-kimbundu_sentence-pairs.fulah-igbo_sentence-pairs
Fulah-Igbo_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Igbo_Sentence-Pairs
Number of Rows: 111376
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-igbo_sentence-pairs.fulah-umbundu_sentence-pairs
Fulah-Umbundu_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Umbundu_Sentence-Pairs
Number of Rows: 30013
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-umbundu_sentence-pairs.fulah-tumbuka_sentence-pairs
Fulah-Tumbuka_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Tumbuka_Sentence-Pairs
Number of Rows: 60296
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-tumbuka_sentence-pairs.fulah-tsonga_sentence-pairs
Fulah-Tsonga_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Tsonga_Sentence-Pairs
Number of Rows: 67239
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-tsonga_sentence-pairs.fulah-lingala_sentence-pairs
Fulah-Lingala_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Lingala_Sentence-Pairs
Number of Rows: 50170
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-lingala_sentence-pairs.fulah-kinyarwanda_sentence-pairs
Fulah-Kinyarwanda_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Kinyarwanda_Sentence-Pairs
Number of Rows: 220054… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-kinyarwanda_sentence-pairs.fulah-xhosa_sentence-pairs
Fulah-Xhosa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Xhosa_Sentence-Pairs
Number of Rows: 167072
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-xhosa_sentence-pairs.fulah-swati_sentence-pairs
Fulah-Swati_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Swati_Sentence-Pairs
Number of Rows: 30447
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-swati_sentence-pairs.fon-fulah_sentence-pairs
Fon-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fon-Fulah_Sentence-Pairs
Number of Rows: 81491
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fon-fulah_sentence-pairs.akan-fulah_sentence-pairs
Akan-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Akan-Fulah_Sentence-Pairs
Number of Rows: 14234
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-fulah_sentence-pairs.
