CoolFace
13 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01beta3 /GridCorpus_9M_Sudoku_Puzzles_Enriched ╔══════════════════════════════════════════════════════════════════════╗ ║ ║ ║ G R I D C O R P U S ║ ║ ║ ║ "004300209005009001070060043..." ║ ║ │ ║ ║ ▼… See the full description on the dataset page: https://huggingface.co/datasets/beta3/GridCorpus_9M_Sudoku_Puzzles_Enriched.tabularfeature-extraction1M<n<10M1 likes1.3k downloads7mo agoHugging Face02Orion-The-Lab /wooden_window_factory_01_enriched_v2 Real industrial data, AI-ready for Physical AI ORION WWF1 – Certified Sample Pack v2.0 (Enriched) Version Status Sector Pipeline v2.0-Enriched 🟢 Level 3 Certified Industrial-Manufacturing Orion Unified V5.2 🌟 The Evolution: Beyond Anonymization The ORION WWF1 v2.0 Enriched pack represents the professional evolution of our baseline industrial dataset. While previous versions focused on privacy-first anonymization, v2.0 transforms raw video… See the full description on the dataset page: https://huggingface.co/datasets/Orion-The-Lab/wooden_window_factory_01_enriched_v2.imagevideo-classificationn<1K0 likes140 downloads5mo agoHugging Face03Orion-The-Lab /wooden_window_factory_01_enriched_v1.5 Real industrial data, AI-ready for Physical AI ORION WWF1 – Enriched Sample Pack v1.5 (Physical AI Edition) 🌟 The Evolution: Beyond Anonymization The ORION WWF1 v1.5 Enriched pack is the professional evolution of our baseline industrial dataset. While version 1.0 focused on privacy-first anonymization, v1.5 transforms raw video into actionable intelligence. This pack includes 10 representative clips from a high-intensity wood-processing facility, now… See the full description on the dataset page: https://huggingface.co/datasets/Orion-The-Lab/wooden_window_factory_01_enriched_v1.5.tabularvideo-classificationn<1K0 likes44 downloads6mo agoHugging Face04mmmaurer /bigbenchhard-enrichedtabular1K<n<10K0 likes40 downloads11mo agoHugging Face05mmmaurer /MAGE-enriched Info This is a version of the MAGE benchmark enriched with linguistic features. The linguistic features were extracted with elfen. Citation If you use this enriched version of MAGE, please cite @inproceedings{doenmez-maurer-2025-ai, title = "AI Argues Differently: Distinct Argumentative and Linguistic Patterns of LLMs in Persuasive Contexts", author = "Dönmez, Esra and Maurer, Maximilian and Lapesa, Gabriella and Falenska, Agnieszka", year = {2025}, booktitle = "To… See the full description on the dataset page: https://huggingface.co/datasets/mmmaurer/MAGE-enriched.tabulartext-classification100K<n<1M0 likes29 downloads1y agoHugging Face06cloud19 /gelbooru-characters-enriched Gelbooru Characters Enriched This dataset is an enriched, fully-mapped version of Gelbooru character tags. It contains resolved franchise (copyright) associations and core appearance features (core tags) for 263,441 unique characters. Dataset Details The dataset maps the original character list to their corresponding copyrights (franchises) and general core attributes. It was constructed using a multi-stage hybrid extraction pipeline: Regex Extraction: Extracting… See the full description on the dataset page: https://huggingface.co/datasets/cloud19/gelbooru-characters-enriched.tabulartext-classification100K<n<1M0 likes23 downloads3mo agoHugging Face07CALM-Lab-Purdue /EnrichedMeaningDataset Dataset Card for Chinese Degree Expressions for Pragmatic Reasoning (CDE-Prag), an ongoing project about Enriched Meaning. Dataset Summary CDE-Prag is a theory-driven evaluation dataset designed to probe the pragmatic competence of Large Language Models (LLMs) and Vision-Language Models (VLMs). It focuses specifically on manner implicatures and ambiguity detection through the lens of Chinese degree expressions (e.g., Kai gao, which is ambiguous between "Kai is tall" and… See the full description on the dataset page: https://huggingface.co/datasets/CALM-Lab-Purdue/EnrichedMeaningDataset.tabularquestion-answeringn<1K1 likes15 downloads7mo agoHugging Face08mmmaurer /mmlu-pro-enrichedtabular10K<n<100K0 likes14 downloads11mo agoHugging Face09bziemba /review-aspects-enrichedSynthetic ahh dataset For each aspect: 0 = negative, 1 = not mentioned, 2 = positive. For overall sentiment: 0 = negative, 2 = positive tabulartext-classification1K<n<10K0 likes12 downloads10mo agoHugging Face10pythn /pois-enriched-usa Enriched Dataset with our stats tabular100K<n<1M0 likes10 downloads1y agoHugging Face11mmmaurer /enriched-generated-arguments Info This is a version of a generated arguments corpus enriched with linguistic features and argument quality dimensions. The linguistic features were extracted with elfen. The argument quality dimensions were extracte with these adapters. Citation If you use this enriched version of the generated arguments corpus, please cite @inproceedings{doenmez-maurer-2025-ai, title = "AI Argues Differently: Distinct Argumentative and Linguistic Patterns of LLMs in Persuasive… See the full description on the dataset page: https://huggingface.co/datasets/mmmaurer/enriched-generated-arguments.tabulartext-classification10K<n<100K0 likes10 downloads1y agoHugging Face12divyaranibth /restaurant_enrichedtabular10K<n<100K0 likes10 downloads4mo agoHugging Face13San0160 /Books_enrichedtabular10K<n<100K0 likes3 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.