CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01umairinayat /sanad_experimentsimage10K<n<100K0 likes255 downloads3mo agoHugging Face02QinEmPeRoR93 /Sanad-ar-dataset0 likes218 downloads12d agoHugging Face03arbml /SANAD Dataset Card for [Dataset Name] Dataset Summary [More Information Needed] Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation Curation Rationale [More Information Needed] Source Data… See the full description on the dataset page: https://huggingface.co/datasets/arbml/SANAD.text100K<n<1M3 likes187 downloads2y agoHugging Face04arbml /Sanadsettext100K<n<1M4 likes89 downloads4y agoHugging Face05freococo /650k_sanadset Sanadset 650K: Data on Hadith Narrators Dataset Description Sanadset is a large-scale dataset containing over 650,986 Hadith records collected from 926 historical Arabic books. This dataset was created to assist in the computational analysis of Islamic Hadiths, specifically focusing on the chain of narrators (Sanad) and the content (Matn). It allows researchers to apply Machine Learning and NLP techniques to tasks such as: Classifying Hadiths (Strong/Weak). Analyzing… See the full description on the dataset page: https://huggingface.co/datasets/freococo/650k_sanadset.tabulartext-classification100K<n<1M0 likes78 downloads7mo agoHugging Face06khalidalt /SANAD Dataset Card for SANAD Dataset Summary SANAD Dataset is a large collection of Arabic news articles that can be used in different Arabic NLP tasks such as Text Classification and Word Embedding. The articles were collected using Python scripts written specifically for three popular news websites: AlKhaleej, AlArabiya and Akhbarona. All datasets have seven categories [Culture, Finance, Medical, Politics, Religion, Sports and Tech], except AlArabiya which doesn’t have… See the full description on the dataset page: https://huggingface.co/datasets/khalidalt/SANAD.text100K<n<1M0 likes56 downloads4y agoHugging Face07Abhishekq10 /sanad-fulltext100K<n<1M0 likes53 downloads4y agoHugging Face08Mouwiya /SANAD Arabic News Articles Dataset About Dataset Context SANAD Dataset is a large collection of Arabic news articles that can be used in different Arabic NLP tasks such as Text Classification and Word Embedding. The articles were collected using Python scripts written specifically for three popular news websites: AlKhaleej, AlArabiya and Akhbarona. All datasets have seven categories [Culture, Finance, Medical, Politics, Religion, Sports and Tech], except AlArabiya which doesn’t… See the full description on the dataset page: https://huggingface.co/datasets/Mouwiya/SANAD.texttext-classification10K<n<100K1 likes49 downloads2y agoHugging Face09ahadda5 /sanadtext100K<n<1M0 likes37 downloads4y agoHugging Face10Kei-Sanada /optuna-logs-task-160 likes34 downloads11mo agoHugging Face11Efficient-Large-Model /sana_data_public0 likes33 downloads2y agoHugging Face12sanadf234 /Heart-Disease-Prediction-datasettabular10K<n<100K1 likes32 downloads10mo agoHugging Face13sanad /semrel0 likes20 downloads3y agoHugging Face14sanad /imdbstext1K<n<10K0 likes11 downloads2y agoHugging Face15Afiqa /arabic-SANAD-5k-sampletext1K<n<10K0 likes11 downloads1y agoHugging Face16Kei-Sanada /optuna-logs-task-17tabularn<1K0 likes11 downloads9mo agoHugging Face17amnamiraj /sanad_experimentsimage10K<n<100K1 likes10 downloads3mo agoHugging Face18Kei-Sanada /optuna-logs-task-22tabularn<1K0 likes8 downloads3mo agoHugging Face19CUTD /sanad_dftext10K<n<100K0 likes5 downloads2y agoHugging Face20Kei-Sanada /optuna-logs-task-210 likes5 downloads5mo agoHugging Face21bigscience-data /roots_ar_sanadgatedROOTS Subset: roots_ar_sanad sanad Dataset uid: sanad Description Homepage Licensing Speaker Locations Sizes 0.1312 % of total 1.2094 % of ar BigScience processing steps Filters applied to: ar dedup_document dedup_template_soft filter_remove_empty_docs remove_html_spans_sanad filter_small_docs_bytes_300 text100K<n<1M0 likes4 downloads4y agoHugging Face22Afiqa /arabic-SANAD-religion-sampletext1K<n<10K0 likes4 downloads1y agoHugging Face23SANAD-GraduationProject2026 /ArabicEmpatheticDialogues-with-Qurantext10K<n<100K0 likes4 downloads2mo agoHugging Face24sanadf234 /FAQs-for-SMEstext10K<n<100K0 likes3 downloads10mo agoHugging Face25Kei-Sanada /optuna-logs-task-19tabularn<1K0 likes3 downloads8mo agoHugging Face26lioradCo /sana_dataset Data Collection Persian domains were collected from the public web and manually reviewed by human annotators. Each domain was categorized based on its primary topic. After the annotation phase, the domains were crawled using a high-speed distributed crawler. The downloaded webpages were processed in parallel by multiple extraction pipelines. The crawling system stores metadata related to each page, including crawl timestamps, parent links, domain information, and page depth.… See the full description on the dataset page: https://huggingface.co/datasets/lioradCo/sana_dataset.textn<1K0 likes3 downloads4mo agoHugging Face27SANAD-GraduationProject2026 /ELQV0 likes3 downloads2mo agoHugging Face28sanadf234 /SMEs-datasettabular10K<n<100K0 likes2 downloads10mo agoHugging Face29Kei-Sanada /optuna-logs-task-200 likes2 downloads7mo agoHugging Face30SanaDG /medquad0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.