CoolFace
27 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01owaiskha9654 /PubMed_MultiLabel_Text_Classification_Dataset_MeSHThis dataset consists of a approx 50k collection of research articles from PubMed repository. Originally these documents are manually annotated by Biomedical Experts with their MeSH labels and each articles are described in terms of 10-15 MeSH labels. In this Dataset we have huge numbers of labels present as a MeSH major which is raising the issue of extremely large output space and severe label sparsity issues. To solve this Issue Dataset has been Processed and mapped to its root as Described… See the full description on the dataset page: https://huggingface.co/datasets/owaiskha9654/PubMed_MultiLabel_Text_Classification_Dataset_MeSH.tabulartext-classification10K<n<100K27 likes210 downloads4y agoHugging Face02jakeazcona /short-text-multi-labeled-emotion-classificationtabular10K<n<100K2 likes61 downloads5y agoHugging Face03syke9p3 /multilabel-tagalog-hate-speechtabular1K<n<10K0 likes36 downloads2y agoHugging Face04chillies /course-review-multilabel-sentiment-analysistabular1K<n<10K0 likes32 downloads2y agoHugging Face05Nasim435 /Multi-label-Prompt-Dataset Multi-label Prompt Dataset A multi-label prompt classification corpus designed for training lightweight, CPU-efficient machine learning models (such as CatBoost, LightGBM, and FastText) for prompt complexity estimation, task intent classification, output token length forecasting, and dynamic LLM routing. Dataset Summary Total Unique Samples: 1,859 deduplicated prompts Number of Classes: 23 multi-label tags across 4 semantic dimensions Language: English (en)… See the full description on the dataset page: https://huggingface.co/datasets/Nasim435/Multi-label-Prompt-Dataset.texttext-classification1K<n<10K0 likes27 downloads1mo agoHugging Face06victoriadreis /TuPY_dataset_multilabel Portuguese Hate Speech Dataset (TuPy) The Portuguese hate speech dataset (TuPy) is an annotated corpus designed to facilitate the development of advanced hate speech detection models using machine learning (ML) and natural language processing (NLP) techniques. TuPy is formed by 10000 thousand unpublished annotated tweets collected in 2023. This repository is organized as follows: root. ├── annotations : classification given by annotators ├── raw corpus : dataset before… See the full description on the dataset page: https://huggingface.co/datasets/victoriadreis/TuPY_dataset_multilabel.tabulartext-classification10K<n<100K3 likes21 downloads3y agoHugging Face07sanjeettoosi /multi-label-text-classificationtext1K<n<10K0 likes20 downloads2y agoHugging Face08hojzas /setfit-proj8-multilabel_2tabularn<1K0 likes17 downloads3y agoHugging Face09nhantran4425 /vnexpress-news-multilabel-2025 VnExpress News Multi-label Dataset 2025 Giới thiệu Bộ dữ liệu ~18,500 bài báo từ VnExpress.net, được gán nhãn đa nhãn với 88 chủ đề. Phù hợp cho bài toán phân loại văn bản tiếng Việt (Vietnamese text classification). Thống kê Train: 14,860 bài Test: 3,715 bài Số nhãn: 88 Ngôn ngữ: Tiếng Việt Tiền xử lý Word segmentation: underthesea Stopwords removal One-hot encoding nhãn Cấu trúc content_final: title×3 + description×2 + content (đã… See the full description on the dataset page: https://huggingface.co/datasets/nhantran4425/vnexpress-news-multilabel-2025.tabulartext-classification10K<n<100K0 likes17 downloads5mo agoHugging Face10csolheim /risk_sig_train_multilabel_OPRtext1K<n<10K0 likes15 downloads3y agoHugging Face11sumaiya-afroze /Multi-Label_Bangla_Hate_Speech_Datareadme_text = """ Bangla Hate Speech Extended Dataset 📖 Overview This dataset is an expanded version of the original Bengali Hate Speech Dataset created by Hriteshwar Talukder and Md Saiful Islam. The original dataset provided a strong foundation for hate speech detection in the Bengali language. In this extended version, the dataset has been: Expanded in size with ~5000 additional Bengali social media comments. Reclassified with fine-grained categories… See the full description on the dataset page: https://huggingface.co/datasets/sumaiya-afroze/Multi-Label_Bangla_Hate_Speech_Data.tabulartext-classification10K<n<100K0 likes15 downloads11mo agoHugging Face12raselmeya2194 /Bangla-Multi-Label-Text_AnalysisThis dataset is a consolidated collection of four public Bengali text datasets curated for sentiment analysis, toxic comment classification and bengali news classification. It consists of Bengali text comments annotated with multiple categories, covering a wide range of sentiment and content-based labels. Aimed at advancing research in Bengali language processing, this dataset is particularly suited for tasks like sentiment analysis, hate speech detection, and contextual comment… See the full description on the dataset page: https://huggingface.co/datasets/raselmeya2194/Bangla-Multi-Label-Text_Analysis.texttext-classification10K<n<100K0 likes14 downloads2y agoHugging Face13csolheim /risk_sig_train_multilabel_FIN_25ktext10K<n<100K0 likes13 downloads3y agoHugging Face14hojzas /setfit-proj8-multilabel_2_validationtabularn<1K0 likes11 downloads3y agoHugging Face15Sharath45 /mentalhealth_multilabel_classification Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [Sharath Ragav] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More Information… See the full description on the dataset page: https://huggingface.co/datasets/Sharath45/mentalhealth_multilabel_classification.tabular10K<n<100K0 likes10 downloads2y agoHugging Face16irfan-ahmad /AURA-Multi_Label_Classification AURA-Classification (Multi-Label Version) Dataset Description The AURA (App User Review in Arabic) Classification dataset is a collection of 2,900 Arabic-language app reviews collected from various mobile applications. This dataset is designed for multi-label text classification, where each review can belong to multiple classes simultaneously. Each review in the dataset was independently annotated by five different annotators. To construct the multi-label version of the… See the full description on the dataset page: https://huggingface.co/datasets/irfan-ahmad/AURA-Multi_Label_Classification.text1K<n<10K1 likes10 downloads9mo agoHugging Face17Alok64 /multilabel_finance_email_inquiriestabular1K<n<10K0 likes9 downloads2y agoHugging Face18hojzas /proj8-multilabeltabularn<1K0 likes7 downloads3y agoHugging Face19bsen26 /eyeR-classification-multi-label-category2tabularn<1K0 likes7 downloads2y agoHugging Face20Ahasan1999 /Multilabel_Emotiontabularn<1K0 likes5 downloads3y agoHugging Face21hojzas /proj8-multilabel-validationtabularn<1K0 likes4 downloads3y agoHugging Face22sg247 /multilabel-classificationtabular10K<n<100K0 likes3 downloads3y agoHugging Face23bsen26 /eyeR-classification-multi-label-category1tabular1K<n<10K0 likes3 downloads2y agoHugging Face24praisethefool /multilabel-sociotechnical_imaginaries-2025_05_06tabularn<1K0 likes2 downloads1y agoHugging Face25KIT-RoboInfo /multi-label-review-1000tabular1K<n<10K0 likes1 downloads2y agoHugging Face26sayedyounes /stack_multilabel_subsettabular1K<n<10K0 likes1 downloads10mo agoHugging Face27KotaroOmote /rg-7wildlife-multilabel-v1image1K<n<10K0 likes1 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.