CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01michsethowusu /umbundu-emotions-corpus Umbundu Emotion Analysis Corpus Dataset Description This dataset contains emotion-labeled text data in Umbundu for emotion classification (joy, sadness, anger, fear, surprise, disgust, neutral). Emotions were extracted and processed from the English meanings of the sentences using the model j-hartmann/emotion-english-distilroberta-base. The dataset is part of a larger collection of African language emotion analysis resources. Dataset Statistics Total samples:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/umbundu-emotions-corpus.texttext-classification10K<n<100K0 likes34 downloads1y agoHugging Face02michsethowusu /umbundu-sentiments-corpus Umbundu Sentiment Corpus Dataset Description This dataset contains sentiment-labeled text data in Umbundu for binary sentiment classification (Positive/Negative). Sentiments are extracted and processed from the English meanings of the sentences using DistilBERT for sentiment classification. The dataset is part of a larger collection of African language sentiment analysis resources. Dataset Statistics Total samples: 83,350 Positive sentiment: 48940 (58.7%)… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/umbundu-sentiments-corpus.texttext-classification10K<n<100K0 likes31 downloads1y agoHugging Face03michsethowusu /akan-umbundu_sentence-pairs Akan-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Akan-Umbundu_Sentence-Pairs Number of Rows: 22651 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-umbundu_sentence-pairs.text10K<n<100K0 likes27 downloads1y agoHugging Face04michsethowusu /swahili-umbundu_sentence-pairs Swahili-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Swahili-Umbundu_Sentence-Pairs Number of Rows: 276674 Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/swahili-umbundu_sentence-pairs.text100K<n<1M0 likes23 downloads1y agoHugging Face05michsethowusu /english-umbundu_sentence-pairs_mt560 English-Umbundu Parallel Dataset This dataset contains parallel sentences in English and Umbundu (Angola). Dataset Information Language Pair: English ↔ Umbundu Language Code: umb Country: Angola Original Source: OPUS MT560 Dataset Dataset Structure The dataset contains parallel sentences that can be used for: Machine translation training Cross-lingual NLP tasks Language model fine-tuning Citation If you use this dataset, please cite the citation… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/english-umbundu_sentence-pairs_mt560.text100K<n<1M0 likes22 downloads1y agoHugging Face06michsethowusu /kongo-umbundu_sentence-pairs Kongo-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kongo-Umbundu_Sentence-Pairs Number of Rows: 58711 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kongo-umbundu_sentence-pairs.text10K<n<100K0 likes20 downloads1y agoHugging Face07michsethowusu /tsonga-umbundu_sentence-pairs Tsonga-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Tsonga-Umbundu_Sentence-Pairs Number of Rows: 132324 Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/tsonga-umbundu_sentence-pairs.text100K<n<1M0 likes19 downloads1y agoHugging Face08michsethowusu /shona-umbundu_sentence-pairs Shona-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Shona-Umbundu_Sentence-Pairs Number of Rows: 153467 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/shona-umbundu_sentence-pairs.text100K<n<1M0 likes19 downloads1y agoHugging Face09michsethowusu /oromo-umbundu_sentence-pairs Oromo-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Oromo-Umbundu_Sentence-Pairs Number of Rows: 48783 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/oromo-umbundu_sentence-pairs.text10K<n<100K0 likes16 downloads1y agoHugging Face10michsethowusu /french-umbundu_sentence-pairstext100K<n<1M0 likes16 downloads1y agoHugging Face11michsethowusu /nuer-umbundu_sentence-pairs Nuer-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Nuer-Umbundu_Sentence-Pairs Number of Rows: 9607 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/nuer-umbundu_sentence-pairs.text1K<n<10K0 likes15 downloads1y agoHugging Face12michsethowusu /tumbuka-umbundu_sentence-pairs Tumbuka-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Tumbuka-Umbundu_Sentence-Pairs Number of Rows: 99125 Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/tumbuka-umbundu_sentence-pairs.text10K<n<100K0 likes14 downloads1y agoHugging Face13michsethowusu /somali-umbundu_sentence-pairs Somali-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Somali-Umbundu_Sentence-Pairs Number of Rows: 96450 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/somali-umbundu_sentence-pairs.text10K<n<100K0 likes14 downloads1y agoHugging Face14michsethowusu /kinyarwanda-umbundu_sentence-pairs Kinyarwanda-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kinyarwanda-Umbundu_Sentence-Pairs Number of Rows:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kinyarwanda-umbundu_sentence-pairs.text100K<n<1M0 likes14 downloads1y agoHugging Face15michsethowusu /dinka-umbundu_sentence-pairs Dinka-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Dinka-Umbundu_Sentence-Pairs Number of Rows: 11337 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/dinka-umbundu_sentence-pairs.text10K<n<100K0 likes14 downloads1y agoHugging Face16michsethowusu /lingala-umbundu_sentence-pairs Lingala-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Lingala-Umbundu_Sentence-Pairs Number of Rows: 86575 Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/lingala-umbundu_sentence-pairs.text10K<n<100K0 likes13 downloads1y agoHugging Face17michsethowusu /kikuyu-umbundu_sentence-pairs Kikuyu-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kikuyu-Umbundu_Sentence-Pairs Number of Rows: 29453 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kikuyu-umbundu_sentence-pairs.text10K<n<100K0 likes13 downloads1y agoHugging Face18michsethowusu /ganda-umbundu_sentence-pairs Ganda-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Ganda-Umbundu_Sentence-Pairs Number of Rows: 76417 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/ganda-umbundu_sentence-pairs.text10K<n<100K0 likes13 downloads1y agoHugging Face19michsethowusu /tswana-umbundu_sentence-pairs Tswana-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Tswana-Umbundu_Sentence-Pairs Number of Rows: 109541 Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/tswana-umbundu_sentence-pairs.text100K<n<1M0 likes12 downloads1y agoHugging Face20michsethowusu /english-umbundu_sentence-pairstext100K<n<1M0 likes11 downloads1y agoHugging Face21michsethowusu /kimbundu-umbundu_sentence-pairs Kimbundu-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kimbundu-Umbundu_Sentence-Pairs Number of Rows: 47912… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kimbundu-umbundu_sentence-pairs.text10K<n<100K0 likes10 downloads1y agoHugging Face22michsethowusu /kamba-umbundu_sentence-pairs Kamba-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kamba-Umbundu_Sentence-Pairs Number of Rows: 41134 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kamba-umbundu_sentence-pairs.text10K<n<100K0 likes10 downloads1y agoHugging Face23michsethowusu /fon-umbundu_sentence-pairs Fon-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Fon-Umbundu_Sentence-Pairs Number of Rows: 56632 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fon-umbundu_sentence-pairs.text10K<n<100K0 likes10 downloads1y agoHugging Face24michsethowusu /ewe-umbundu_sentence-pairs Ewe-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Ewe-Umbundu_Sentence-Pairs Number of Rows: 106645 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/ewe-umbundu_sentence-pairs.text100K<n<1M0 likes10 downloads1y agoHugging Face25michsethowusu /bambara-umbundu_sentence-pairs Bambara-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Bambara-Umbundu_Sentence-Pairs Number of Rows: 20038 Number… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bambara-umbundu_sentence-pairs.text10K<n<100K0 likes10 downloads1y agoHugging Face26michsethowusu /tigrinya-umbundu_sentence-pairs Tigrinya-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Tigrinya-Umbundu_Sentence-Pairs Number of Rows: 66683… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/tigrinya-umbundu_sentence-pairs.text10K<n<100K0 likes9 downloads1y agoHugging Face27michsethowusu /igbo-umbundu_sentence-pairs Igbo-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Igbo-Umbundu_Sentence-Pairs Number of Rows: 58532 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/igbo-umbundu_sentence-pairs.text10K<n<100K0 likes9 downloads1y agoHugging Face28michsethowusu /fulah-umbundu_sentence-pairs Fulah-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Fulah-Umbundu_Sentence-Pairs Number of Rows: 30013 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-umbundu_sentence-pairs.text10K<n<100K0 likes9 downloads1y agoHugging Face29michsethowusu /dyula-umbundu_sentence-pairs Dyula-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Dyula-Umbundu_Sentence-Pairs Number of Rows: 37912 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/dyula-umbundu_sentence-pairs.text10K<n<100K0 likes9 downloads1y agoHugging Face30michsethowusu /chichewa-umbundu_sentence-pairs Chichewa-Umbundu_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Chichewa-Umbundu_Sentence-Pairs Number of Rows: 141904… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/chichewa-umbundu_sentence-pairs.text100K<n<1M0 likes9 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.