sitong-fang/MM-DeceptionBench
🎭 MM-DeceptionBench A Multimodal Benchmark for Evaluating Deceptive Behaviors in Vision-Language Models 📖 Overview MM-DeceptionBench is a comprehensive benchmark designed to stress-test Multimodal Large Language Models (MLLMs) for strategic deception in visually grounded contexts. It captures nuanced deceptive behaviors that emerge when models interact with images and text, spanning diverse real-world scenarios. ✨ Key… See the full description on the dataset page: https://huggingface.co/datasets/sitong-fang/MM-DeceptionBench.
025
