CoolFace
Datasetpublic

Rakancorle1/hans-sft-4k

Hans-SFT-4K · SFT recipe for the audio-visual Clever Hans Supervised fine-tuning (SFT) data accompanying the paper When Vision Speaks for Sound. Like the original Clever Hans — the horse that looked like he could do arithmetic but was actually reading his trainer's body language — video-capable MLLMs often look like they can hear: they answer audio questions by reading visual cues and never verifying the audio stream. Hans-SFT-4K is the 3,834-sample SFT mix that teaches models… See the full description on the dataset page: https://huggingface.co/datasets/Rakancorle1/hans-sft-4k.

sourceHugging Facecc-by-nc-4.0updated 4mo agoView on Hugging Face
1likes27downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
Rakancorle1/hans-sft-4k · CoolFace