CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01yatin-superintelligence /Edge-Agent-Reasoning-WebSearch-260K Edge Agent Reasoning WebSearch 260K Abstract The Edge-Agent-Reasoning-WebSearch-260K dataset is a massive, synthetically expert-engineered corpus of over 700 Million tokens, designed to train small, local models (SLMs) and edge-deployed agents in advanced problem deconstruction and self-aware reasoning. Rather than training a model to execute instructions directly—which often leads to hallucinations when context is missing—this dataset trains a model to act as a… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Edge-Agent-Reasoning-WebSearch-260K.texttext-generation100K<n<1M53 likes5.7k downloads7mo agoHugging Face02yatin-superintelligence /White-Hat-Security-Agent-Prompts-600K White Hat Security Agent Prompts 600K Overview The White-Hat-Security-Agent-Prompts-600K dataset is a practitioner-perspective security prompts corpus of 596,295 richly contextualized queries, designed to represent how real-world defensive security professionals communicate, interrogate, and reason through active threat scenarios. Where most security datasets catalogue CVEs, malware signatures, or CTF write-ups, this collection teaches models to operate from inside the… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/White-Hat-Security-Agent-Prompts-600K.texttext-generation100K<n<1M21 likes1.9k downloads7mo agoHugging Face03yatin-superintelligence /Creative-Professionals-Agentic-Tasks-1M Creative Professionals Agentic Tasks (1M) Abstract A massive-scale, high-fidelity synthetic task dataset comprising 1,070,917 agentic command operations across 36 creative, technical, and engineering software environments. This dataset is engineered exclusively to stress-test, evaluate, and fine-tune multimodal AI agents designed for Agent Environment operation, complex software interaction, and multi-step reasoning within deep software infrastructures.… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Creative-Professionals-Agentic-Tasks-1M.tabulartext-generation1M<n<10M29 likes1.2k downloads7mo agoHugging Face04CyberMax-tools /superintelligence-search-questions Keyfern: what people search about superintelligence (Sept 2026) Need it fresh, filtered or via API? This free file is a snapshot (search questions at the last refresh), last updated 2026-09-25. Keyfern on Apify ($0.001 per keyword idea): gets keyword ideas and questions for your own seeds from 5 search engines. Swellmeter on Apify ($0.003 per keyword analyzed): checks Google Trends direction for your own keywords with a plain rising/falling verdict, daily if scheduled.… See the full description on the dataset page: https://huggingface.co/datasets/CyberMax-tools/superintelligence-search-questions.tabulartext-classification1K<n<10K0 likes69 downloads16h agoHugging Face05CyberMax-tools /superintelligence-ban-bill-coverage Policywren: the superintelligence ban bill in news coverage and search interest (Sept 2026) Need it fresh, filtered or via API? This free file is a snapshot (coverage at the last refresh), last updated 2026-09-25. Swellmeter on Apify ($0.003 per keyword analyzed): checks Google Trends direction for your own keywords with a plain rising/falling verdict, daily if scheduled. Get an email when this dataset updates: free, double opt-in, unsubscribe any time. In September… See the full description on the dataset page: https://huggingface.co/datasets/CyberMax-tools/superintelligence-ban-bill-coverage.tabulartext-classificationn<1K0 likes67 downloads16h agoHugging Face06yatin-superintelligence /Adversarial-Agent-Intent-Safety-Analysis-240Kgated Adversarial Agent Intent Safety Analysis 240K Abstract The Adversarial-Agent-Intent-Safety-Analysis-240K is a deterministically structured dataset featuring 242,454 context-rich adversarial prompts and safety evaluations. Engineered strictly for training frontier command-and-control models, guardrail classifiers, and red-teaming agents, it encourages models to parse multi-layered intention across 126 critical risk vectors. This design trains models to decouple the surface… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Adversarial-Agent-Intent-Safety-Analysis-240K.texttext-classification100K<n<1M12 likes58 downloads7mo agoHugging Face07jedanderson /environmental-superintelligence-corpus Environmental Superintelligence Corpus A machine-readable export of the published writing of Jed Anderson (ORCID 0009-0003-1807-2459) on environmental superintelligence, information physics, and faith-integrated first-principles thinking. The corpus is built around one thesis — "Bits Protect Its": information accumulates causal sovereignty over matter and energy, so directing the world by knowledge is, as a matter of physical law, vastly cheaper than directing it by force. These… See the full description on the dataset page: https://huggingface.co/datasets/jedanderson/environmental-superintelligence-corpus.textn<1K0 likes21 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.