datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
M3Kang
M3Kang: A Multilingual Multimodal Mathematical Reasoning Dataset from Kangaroo Problems
Introduction
Despite state-of-the-art vision-language models (VLMs) have demonstrated strong reasoning capabilities, their performance in multilingual mathematical reasoning remains underexplored. To bridge this gap, we introduce M3Kang, the first massively multilingual, multimodal mathematical reasoning dataset for VLMs. It is derived from the Kangaroo Math Competition, the world’s… See the full description on the dataset page: https://huggingface.co/datasets/qualcomm/M3Kang.csd100
Introduction
Disentangling content and style from a single image, known as content-style decomposition (CSD), enables recontextualization of extracted content and stylization of extracted styles, offering greater creative flexibility in visual synthesis. While existing datasets focus on either style transfer or content preservation, they do not fully meet the requirements for evaluating CSD, prompting us to introduce CSD-100, a dataset of 100 images designed specifically for this… See the full description on the dataset page: https://huggingface.co/datasets/qualcomm/csd100.candv3a4realadsim-baseline-v3-an-hrr-i40krealadsim-baseline-v4-dr-25krealadsim-baseline-v4-dr-an-hrr-i40k
