sitong-fang/MM-DeceptionBench
🎠MM-DeceptionBench A Multimodal Benchmark for Evaluating Deceptive Behaviors in Vision-Language Models 📖 Overview MM-DeceptionBench is a comprehensive benchmark designed to stress-test Multimodal Large Language Models (MLLMs) for strategic deception in visually grounded contexts. It captures nuanced deceptive behaviors that emerge when models interact with images and text, spanning diverse real-world scenarios. ✨ Key… See the full description on the dataset page: https://huggingface.co/datasets/sitong-fang/MM-DeceptionBench.
This repository belongs to sitong-fang on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
