datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
HawkEye-IT
Download Video
Please download the original videos from the provided links:
VideoChat: Based on InternVid, we created additional instruction data and used GPT-4 to condense the existing data.
VideoChatGPT: The original caption data was converted into conversation data based on the same VideoIDs.
Kinetics-710 & SthSthV2: Option candidates were generated from UMTtop-20 predictions.
NExTQA: Typos in the original sentences were corrected.
CLEVRER: For single-option multiple-choice QAs… See the full description on the dataset page: https://huggingface.co/datasets/wangyueqian/HawkEye-IT.HawkBenchpm-agi-benchmark
PM-AGI Benchmark v2 🎯
The open-source LLM reasoning benchmark for Performance Marketing.
Developed by hawky.ai — evaluating how well LLMs reason about real-world Meta Ads and Google Ads scenarios. v2 (494 questions) is built to surface the gap between knowledge recall and genuine reasoning.
Dataset Summary (v2)
PM-AGI v2 contains 494 expert-crafted questions across 4 categories and 5 reasoning types:
Category
Questions
Focus
Meta Ads
227
Campaign structure… See the full description on the dataset page: https://huggingface.co/datasets/Hawky-ai/pm-agi-benchmark.
