CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Bambolbb /cits4012_A1_2026_medical_abstractstext10K<n<100K0 likes99 downloads18d agoHugging Face02LLM4APR /A11YBench A11YBench A Benchmark for Web Accessibility Repair 😃Dataset Summary A11YBench consists of 60 real-world web projects, encompassing 147 web pages and 8,886 accessibility violations detected by the IBM Accessibility Checker using Check Rule 2025.09.03. The projects vary substantially in size, from 123 to 43,198 source files and from 3,610 to 1,555,532 lines of code, covering both lightweight documentation sites and large production-grade applications. This scale ensures… See the full description on the dataset page: https://huggingface.co/datasets/LLM4APR/A11YBench.tabularn<1K0 likes47 downloads8mo agoHugging Face03lopi-hfsec-a1 /lopi-ds-conv-a1textn<1K0 likes44 downloads14d agoHugging Face04tarosato /mad-blow-a17d80 mad-blow-a17d80 Synthetic sensors test data: 53 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/tarosato/mad-blow-a17d80.tabularn<1K0 likes31 downloads12d agoHugging Face05yumiko89 /strong-government-a16a25 strong-government-a16a25 Synthetic weather test data: 39 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/yumiko89/strong-government-a16a25.tabularn<1K0 likes29 downloads12d agoHugging Face06TianfuXinqu /pwc747_a10865__paper__P02__2023__high__llm_safety Northwind Support Tickets Archive A derived dataset combining service interaction logs with survey responses for support ticket analysis. Upstream Sources This dataset is derived from the following upstream source datasets: Northwind Service Interaction Logs (TianfuXinqu/pwc747_a10865__paper__P05__2022__high__llm_safety) Northwind Customer Survey Responses (TianfuXinqu/pwc747_a10865__paper__P06__2018__low__federated_learning) Commercial Use… See the full description on the dataset page: https://huggingface.co/datasets/TianfuXinqu/pwc747_a10865__paper__P02__2023__high__llm_safety.textn<1K0 likes28 downloads1mo agoHugging Face07Azure-Ivan98 /eastern-reputation-a1327d eastern-reputation-a1327d Synthetic products test data: 50 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at… See the full description on the dataset page: https://huggingface.co/datasets/Azure-Ivan98/eastern-reputation-a1327d.tabularn<1K0 likes28 downloads12d agoHugging Face08TianfuXinqu /pwc747_a10865__paper__P01__2024__high__federated_learning Northwind Product Reviews Corpus A derived dataset combining customer profiles with store reviews to support product review analysis. Upstream Sources This dataset is derived from the following upstream source datasets: Northwind Customer Profiles (TianfuXinqu/pwc747_a10865__paper__P03__2021__low__federated_learning) Northwind Store Reviews (TianfuXinqu/pwc747_a10865__paper__P04__2019__high__model_compression) Commercial Use Commercial Use:… See the full description on the dataset page: https://huggingface.co/datasets/TianfuXinqu/pwc747_a10865__paper__P01__2024__high__federated_learning.textn<1K0 likes25 downloads1mo agoHugging Face09siberiainstitute-a11y /tourism-package-prediction-datatabular1K<n<10K0 likes21 downloads2d agoHugging Face10SeanSha30 /swedish-pre-a1-scenario-classifier-dataset Swedish Pre-A1 Scenario Classification Dataset This dataset contains 150 short Swedish learner sentences for text classification. It is designed for absolute beginner / pre-A1 learners and aligned with beginner Swedish lecture themes. Labels food_shop family_school health_places transport home_places social_intro Each label has 25 examples. Columns id: unique example id text: Swedish learner sentence used as model input english: English translation chinese:… See the full description on the dataset page: https://huggingface.co/datasets/SeanSha30/swedish-pre-a1-scenario-classifier-dataset.texttext-classificationn<1K0 likes19 downloads4mo agoHugging Face11qicheng9481-a11y /uspto-patent-datatabular100K<n<1M1 likes16 downloads2mo agoHugging Face12BekzatK /shyrai-a1c-quality-control Shyrai A1c Quality Control Dataset Description This dataset contains real-world quality control (QC) data collected during the production and testing of the Shyrai A1c glycated hemoglobin analyzer, used for diabetes diagnostics. The dataset is designed to support research in: AI-driven quality management systems (QMS) anomaly detection predictive quality analytics medical device manufacturing Dataset Structure The dataset includes the following types of… See the full description on the dataset page: https://huggingface.co/datasets/BekzatK/shyrai-a1c-quality-control.tabulartabular-classificationn<1K0 likes12 downloads5mo agoHugging Face13lopi-hfsec-a1 /dsv-gated-chain-turw0z1gatedtext1K<n<10K0 likes12 downloads13d agoHugging Face14a1b8h04i /Hinglish-Everyday-Conversations-1M Dataset Card for Hinglish Everyday Conversations Dataset A synthetically created Hinglish-based dataset of 2 columns where every row represents a unique conversation between 2 people in Hinglish about Everyday Life Topics. Use Model Access the model made using this dataset: Tiny-Hinglish-Chat-21M For more information about this model, its training process, or related resources, you can check the GitHub repository Tiny-Hinglish-Chat-21M-Scripts. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/a1b8h04i/Hinglish-Everyday-Conversations-1M.texttext-generation1M<n<10M0 likes11 downloads4mo agoHugging Face15A1221chal /SanskritDatasettext100K<n<1M1 likes10 downloads5mo agoHugging Face16lopi-hfsec-a1 /lopi-ds-gated-a1gatedtextn<1K0 likes10 downloads14d agoHugging Face17SoorajK1 /testing_01-e61c8c4f-3dcd-4cd8-a1db-eadb7f5c3394tabularn<1K0 likes7 downloads3y agoHugging Face18dz-data-ai /a10_command_scripttextn<1K0 likes5 downloads3y agoHugging Face19a10144 /sa_data_origin감성대화 말뭉치와 한국어 단발성 대화 병합 데이터셋 label : "불안", "분노", "상처", "슬픔", "당황", "기쁨", "놀람" text10K<n<100K0 likes2 downloads1y agoHugging Face20FIRSTACCOUNT69 /sqli-a1-envtextn<1K0 likes2 downloads6mo agoHugging Face21DQLink /a1text1K<n<10K0 likes1 downloads1y agoHugging Face22lopi-hfsec-a1 /auto22t5dmuceya5ax tabularn<1K0 likes10h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.