fulah
Datasets
All datasets matching “fulah”fulah-hausa_sentence-pairs
Fulah-Hausa_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Hausa_Sentence-Pairs
Number of Rows: 269337
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-hausa_sentence-pairs.fulah-twi_sentence-pairs
Fulah-Twi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Twi_Sentence-Pairs
Number of Rows: 77858
Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-twi_sentence-pairs.fulah-emotions-corpus
Fulah Emotion Analysis Corpus
Dataset Description
This dataset contains emotion-labeled text data in Fulah for emotion classification (joy, sadness, anger, fear, surprise, disgust, neutral). Emotions were extracted and processed from the English meanings of the sentences using the model j-hartmann/emotion-english-distilroberta-base. The dataset is part of a larger collection of African language emotion analysis resources.
Dataset Statistics
Total samples: 80… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-emotions-corpus.fulah-rundi_sentence-pairs
Fulah-Rundi_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Fulah-Rundi_Sentence-Pairs
Number of Rows: 80036
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-rundi_sentence-pairs.amharic-fulah_sentence-pairs
Amharic-Fulah_Sentence-Pairs Dataset
This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks.
This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1
Metadata
File Name: Amharic-Fulah_Sentence-Pairs
Number of Rows: 435048
Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/amharic-fulah_sentence-pairs.fulah-sentiments-corpus
Fulah Sentiment Corpus
Dataset Description
This dataset contains sentiment-labeled text data in Fulah for binary sentiment classification (Positive/Negative). Sentiments are extracted and processed from the English meanings of the sentences using DistilBERT for sentiment classification. The dataset is part of a larger collection of African language sentiment analysis resources.
Dataset Statistics
Total samples: 80,686
Positive sentiment: 45880 (56.9%)
Negative… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-sentiments-corpus.
