datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Performance-Marketing-Data
Performance Marketing Expert Dataset
Dataset Description
This dataset contains comprehensive performance marketing knowledge and logical reasoning patterns for Meta (Facebook/Instagram), Google Ads, and TikTok advertising platforms. It's designed for fine-tuning language models to understand brand verticals, performance marketing strategies, and develop reasoning capacity for creating winning ad campaigns.
Dataset Structure
Each example follows an… See the full description on the dataset page: https://huggingface.co/datasets/Sri-Vigneshwar-DJ/Performance-Marketing-Data.clinical-quad-site-performance-signal-drift-oversight-lag-v0.1Clarus Clinical Quad Coupling Site Performance Signal Drift Oversight Lag v0.1
What this dataset isThis dataset tests whether a model can detect site-level performance drift driven by four interacting nodes.
Quad coupling nodes
Enrollment or reporting signal shift
Data capture or documentation gaps
Operational staffing or monitoring lag
Governance pressure such as reviews, incentives, or interim analyses
Input
One site vignette
OutputReturn strict JSON only.
Required output JSON… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-quad-site-performance-signal-drift-oversight-lag-v0.1.leetcode-performance
Dataset card for Leetcode Performance Dataset
texas-teks-ultimate-real-data-enhanced-performance-analytics
Texas TEKS Performance Analytics - Real Data Metrics
📋 Dual Licensing Model - Legal Framework
Based on software-legal-counsel agent recommendation for balancing public educational access with commercial attribution requirements.
⚖️ License Selection Guide
This dataset uses a dual licensing model to maximize educational accessibility while ensuring appropriate commercial attribution:
🎓 Educational/Research Use: CC BY 4.0
Use this license if you are:… See the full description on the dataset page: https://huggingface.co/datasets/RobworksSoftware/texas-teks-ultimate-real-data-enhanced-performance-analytics.performance-marketing-2026-critical
🎯 Performance Marketing 2026 Critical Dataset
A premium, reasoning-dense dataset crafted for the 2026 Digital Advertising Landscape. This dataset is specifically designed to train Large Language Models (LLMs) in high-level strategic thinking, campaign diagnosis, and advanced measurement for Meta and Google Ads.
🚀 Why 2026?
The digital marketing world has undergone a radical shift. Deterministic tracking is a thing of the past. AI-driven automation (Meta's ASC and… See the full description on the dataset page: https://huggingface.co/datasets/Sri-Vigneshwar-DJ/performance-marketing-2026-critical.devto-war-story-performance
Dev.to War-Story Performance Dataset
941 articles published on dev.to under the @whoffagents account, spanning April 2026. Includes title, tags, engagement metrics, and reading time.
Why this exists
We run an agentic content pipeline that publishes developer war-stories daily. This dataset captures real performance data across article formats to answer: what titles and tags actually get reactions on dev.to?
Preliminary finding: war-story framing ("I did X and here's what… See the full description on the dataset page: https://huggingface.co/datasets/WH0FF/devto-war-story-performance.Performance_Management_Difficult_Conversations_Practical
Performance Management Difficult Conversations — Practical
This corpus was automatically generated by the Deku Corpus Builder for use in RAG-based AI applications.
Dataset Structure
Each record contains:
text: The content text
source_url: Original source URL
source_title: Title of the source document
source_domain: Domain of the source
license_type: License classification (e.g. public_domain, cc_by, cc_by_sa)
attribution_required: Boolean — True for CC BY / CC BY-SA and… See the full description on the dataset page: https://huggingface.co/datasets/PhillyMac/Performance_Management_Difficult_Conversations_Practical.Performance_Management_Difficult_Conversations_Theory
Performance Management Difficult Conversations — Theory
This corpus was automatically generated by the Deku Corpus Builder for use in RAG-based AI applications.
Dataset Structure
Each record contains:
text: The content text
source_url: Original source URL
source_title: Title of the source document
source_domain: Domain of the source
license_type: License classification (e.g. public_domain, cc_by, cc_by_sa)
attribution_required: Boolean — True for CC BY / CC BY-SA and… See the full description on the dataset page: https://huggingface.co/datasets/PhillyMac/Performance_Management_Difficult_Conversations_Theory.
