datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
remote_sensing_VQA_multilingual
Remote Sensing VQA — Multilingual
A multilingual counterfactual MCQ dataset built from remote sensing / satellite imagery.
Each row contains a satellite image, two captions (original vs counterfactual), and a multiple-choice question probing whether a VLM follows the image or the misleading text.
Languages
Language
Code
Rows
English
en
50
Hindi
hi
50
Urdu
ur
50
Telugu
te
50
Bahasa Indonesia
id
50
Columns
Column
Type… See the full description on the dataset page: https://huggingface.co/datasets/apart-global-south-hack/remote_sensing_VQA_multilingual.counterfactual-pendulum-multilingual
📌 Dataset Summary
When a Vision-Language Model (VLM) is given an image along with a text prompt containing contradictory or misleading information, how does it react? Does it rely on the visual evidence, succumb to textual bias, or honestly abstain when faced with unresolvable conflict?
This dataset adapts the Counterfactual Pendulum scenario across two visual conflict dimensions:
Angular (Angle): Conflict in the pendulum's angle of inclination.
Light: Conflict in the light… See the full description on the dataset page: https://huggingface.co/datasets/apart-global-south-hack/counterfactual-pendulum-multilingual.multilingual-crossmodal-conflict-3D_Objects
Multilingual Cross-Modal Conflict — 3D Objects
A multilingual counterfactual MCQ dataset built from rendered 3D object scenes.
Each row contains a rendered 3D scene image, two captions (original vs counterfactual), and a multiple-choice question probing whether a VLM follows the image or the misleading text.
Languages
Language
Code
Rows
English
en
150
Hindi
hi
150
Telugu
te
150
Bahasa Indonesia
id
150
Columns
Column
Type… See the full description on the dataset page: https://huggingface.co/datasets/apart-global-south-hack/multilingual-crossmodal-conflict-3D_Objects.multilingual-counterfactual
Multilingual Counterfactual
A multilingual counterfactual MCQ dataset built from COCO-Counterfactual.
Each row contains an image, two captions (original vs counterfactual), and a multiple-choice question probing the difference between image content and text description.
Languages
Language
Code
Status
English
en
✅ Done
Hindi
hi
✅ Done
Urdu
ur
✅ Done
Telugu
te
✅ Done
Bahasa Indonesia
id
✅ Done
Columns
Column
Type… See the full description on the dataset page: https://huggingface.co/datasets/apart-global-south-hack/multilingual-counterfactual.
