Tajwer/roman_urdu_negation_blindspot
Roman Urdu Negation Blind Spot Dataset Summary This dataset tests whether Qwen3-4B-Instruct-2507 uses negation correctly when judging Roman Urdu sentences, rather than just being able to translate it. It pairs 12 sentences across two polarities (affirmative and negated), tested on two tasks: English translation and a constrained Yes/No sentiment/factual judgment, with a matched English control set (48 rows total). Built as the technical challenge submission for… See the full description on the dataset page: https://huggingface.co/datasets/Tajwer/roman_urdu_negation_blindspot.
055
