CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01michsethowusu /rundi-sentiments-corpus Rundi Sentiment Corpus Dataset Description This dataset contains sentiment-labeled text data in Rundi for binary sentiment classification (Positive/Negative). Sentiments are extracted and processed from the English meanings of the sentences using DistilBERT for sentiment classification. The dataset is part of a larger collection of African language sentiment analysis resources. Dataset Statistics Total samples: 372,663 Positive sentiment: 209740 (56.3%)… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-sentiments-corpus.texttext-classification100K<n<1M0 likes40 downloads1y agoHugging Face02michsethowusu /english-rundi_sentence-pairs English-Rundi_Sentence-Pairs Dataset This dataset can be used for machine translation, sentence alignment, or other natural language processing tasks. It is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: English-Rundi_Sentence-Pairs File Size: 94265941 bytes Languages: English, English Dataset Description The dataset contains sentence pairs in… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/english-rundi_sentence-pairs.text100K<n<1M0 likes36 downloads1y agoHugging Face03michsethowusu /Code-170k-rundi Dataset Description Code-170k-rundi is a groundbreaking dataset containing 176,999 programming conversations, originally sourced from glaiveai/glaive-code-assistant-v2 and translated into Rundi, making coding education accessible to Rundi speakers. 🌟 Key Features 176,999 high-quality conversations about programming and coding Pure Rundi language - democratizing coding education Multi-turn dialogues covering various programming concepts Diverse topics: algorithms, data… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/Code-170k-rundi.texttext-generation100K<n<1M1 likes27 downloads11mo agoHugging Face04michsethowusu /rundi-tumbuka_sentence-pairs Rundi-Tumbuka_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Tumbuka_Sentence-Pairs Number of Rows: 194527 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-tumbuka_sentence-pairs.text100K<n<1M0 likes25 downloads1y agoHugging Face05michsethowusu /ewe-rundi_sentence-pairs Ewe-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Ewe-Rundi_Sentence-Pairs Number of Rows: 197511 Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/ewe-rundi_sentence-pairs.text100K<n<1M0 likes23 downloads1y agoHugging Face06michsethowusu /english-rundi_sentence-pairs_mt560 English-Rundi Parallel Dataset This dataset contains parallel sentences in English and Rundi (Burundi). Dataset Information Language Pair: English ↔ Rundi Language Code: run Country: Burundi Original Source: OPUS MT560 Dataset Dataset Structure The dataset contains parallel sentences that can be used for: Machine translation training Cross-lingual NLP tasks Language model fine-tuning Citation If you use this dataset, please cite the citation… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/english-rundi_sentence-pairs_mt560.text100K<n<1M0 likes23 downloads1y agoHugging Face07michsethowusu /fulah-rundi_sentence-pairs Fulah-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Fulah-Rundi_Sentence-Pairs Number of Rows: 80036 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fulah-rundi_sentence-pairs.text10K<n<100K0 likes22 downloads1y agoHugging Face08michsethowusu /hausa-rundi_sentence-pairs Hausa-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Hausa-Rundi_Sentence-Pairs Number of Rows: 261557 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/hausa-rundi_sentence-pairs.text100K<n<1M0 likes21 downloads1y agoHugging Face09michsethowusu /rundi-swahili_sentence-pairs Rundi-Swahili_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Swahili_Sentence-Pairs Number of Rows: 622955 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-swahili_sentence-pairs.text100K<n<1M0 likes20 downloads1y agoHugging Face10michsethowusu /kimbundu-rundi_sentence-pairs Kimbundu-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kimbundu-Rundi_Sentence-Pairs Number of Rows: 75678 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kimbundu-rundi_sentence-pairs.text10K<n<100K0 likes20 downloads1y agoHugging Face11michsethowusu /rundi-somali_sentence-pairs Rundi-Somali_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Somali_Sentence-Pairs Number of Rows: 210469 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-somali_sentence-pairs.text100K<n<1M0 likes19 downloads1y agoHugging Face12michsethowusu /rundi-twi_sentence-pairs Rundi-Twi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Twi_Sentence-Pairs Number of Rows: 199589 Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-twi_sentence-pairs.text100K<n<1M0 likes18 downloads1y agoHugging Face13michsethowusu /oromo-rundi_sentence-pairs Oromo-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Oromo-Rundi_Sentence-Pairs Number of Rows: 91889 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/oromo-rundi_sentence-pairs.text10K<n<100K0 likes17 downloads1y agoHugging Face14michsethowusu /akan-rundi_sentence-pairs Akan-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Akan-Rundi_Sentence-Pairs Number of Rows: 47706 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/akan-rundi_sentence-pairs.text10K<n<100K0 likes16 downloads1y agoHugging Face15michsethowusu /rundi-emotions-corpus Rundi Emotion Analysis Corpus Dataset Description This dataset contains emotion-labeled text data in Rundi for emotion classification (joy, sadness, anger, fear, surprise, disgust, neutral). Emotions were extracted and processed from the English meanings of the sentences using the model j-hartmann/emotion-english-distilroberta-base. The dataset is part of a larger collection of African language emotion analysis resources. Dataset Statistics Total samples: 372… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-emotions-corpus.texttext-classification100K<n<1M0 likes16 downloads1y agoHugging Face16michsethowusu /rundi-xhosa_sentence-pairs Rundi-Xhosa_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Xhosa_Sentence-Pairs Number of Rows: 307505 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-xhosa_sentence-pairs.text100K<n<1M0 likes15 downloads1y agoHugging Face17svjack /OnePromptOneStory-RunDiffusion-Juggernaut-X-v10image1K<n<10K1 likes14 downloads2y agoHugging Face18svjack /OnePromptOneStory-RunDiffusion-Juggernaut-XI-v11image1K<n<10K0 likes14 downloads2y agoHugging Face19michsethowusu /french-rundi_sentence-pairstext1M<n<10M0 likes14 downloads1y agoHugging Face20michsethowusu /rundi-yoruba_sentence-pairs Rundi-Yoruba_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Yoruba_Sentence-Pairs Number of Rows: 174926 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-yoruba_sentence-pairs.text100K<n<1M0 likes12 downloads1y agoHugging Face21michsethowusu /kongo-rundi_sentence-pairs Kongo-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kongo-Rundi_Sentence-Pairs Number of Rows: 99480 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kongo-rundi_sentence-pairs.text10K<n<100K0 likes12 downloads1y agoHugging Face22michsethowusu /dyula-rundi_sentence-pairs Dyula-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Dyula-Rundi_Sentence-Pairs Number of Rows: 72663 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/dyula-rundi_sentence-pairs.text10K<n<100K0 likes12 downloads1y agoHugging Face23michsethowusu /bemba-rundi_sentence-pairs Bemba-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Bemba-Rundi_Sentence-Pairs Number of Rows: 208683 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/bemba-rundi_sentence-pairs.text100K<n<1M0 likes12 downloads1y agoHugging Face24michsethowusu /rundi-tsonga_sentence-pairs Rundi-Tsonga_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Tsonga_Sentence-Pairs Number of Rows: 283463 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-tsonga_sentence-pairs.text100K<n<1M0 likes11 downloads1y agoHugging Face25michsethowusu /kikuyu-rundi_sentence-pairs Kikuyu-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Kikuyu-Rundi_Sentence-Pairs Number of Rows: 59660 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/kikuyu-rundi_sentence-pairs.text10K<n<100K0 likes11 downloads1y agoHugging Face26michsethowusu /fon-rundi_sentence-pairs Fon-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Fon-Rundi_Sentence-Pairs Number of Rows: 89002 Number of Columns:… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/fon-rundi_sentence-pairs.text10K<n<100K0 likes11 downloads1y agoHugging Face27michsethowusu /dinka-rundi_sentence-pairs Dinka-Rundi_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Dinka-Rundi_Sentence-Pairs Number of Rows: 24413 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/dinka-rundi_sentence-pairs.text10K<n<100K0 likes11 downloads1y agoHugging Face28svjack /OnePromptOneStory-RunDiffusion-Juggernaut-XI-v11-CCIPimagen<1K0 likes10 downloads2y agoHugging Face29michsethowusu /rundi-shona_sentence-pairs Rundi-Shona_Sentence-Pairs Dataset This dataset contains sentence pairs for African languages along with similarity scores. It can be used for machine translation, sentence alignment, or other natural language processing tasks. This dataset is based on the NLLBv1 dataset, published on OPUS under an open-source initiative led by META. You can find more information here: OPUS - NLLB-v1 Metadata File Name: Rundi-Shona_Sentence-Pairs Number of Rows: 334206 Number of… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/rundi-shona_sentence-pairs.text100K<n<1M0 likes10 downloads1y agoHugging Face30svjack /OnePromptOneStory-RunDiffusion-Juggernaut-X-v10-CCIPimagen<1K0 likes9 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.