datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PuzzleWorld
Dataset Card for PuzzleWorld
PuzzleWorld is a benchmark of 667 real-world puzzlehunt–style problems designed to evaluate open-ended, multimodal reasoning capabilities of AI models. Curated from Puzzled Pint’s Creative Commons–licensed archives (2010–2025), each puzzle combines text, visual, and structured inputs with no explicitly stated instructions. Solvers must first infer the hidden problem structure from ambiguous clues and then execute a multi-step, creative reasoning… See the full description on the dataset page: https://huggingface.co/datasets/hzli1202/PuzzleWorld.visual-puzzlespuzzle-hle-filteredPuzzleVQAPaper | Code | Dataset
About
Large multimodal models extend the impressive capabilities of large language models by integrating multimodal
understanding abilities. However, it is not clear how they can emulate the general intelligence and reasoning ability of
humans. As recognizing patterns and abstracting concepts are key to general intelligence, we introduce PuzzleVQA, a
collection of puzzles based on abstract patterns. With this dataset, we evaluate large multimodal models with… See the full description on the dataset page: https://huggingface.co/datasets/declare-lab/PuzzleVQA.chess-puzzles-images-mini
Dataset Card for Chess Puzzles Images (mini)
This dataset contains 124,999 chess board positions in JPG format, derived from Lichess puzzles. Each image is accompanied by a shortened FEN string, indication for the color to play as, castling and en passant availability, and best moves in standard algebraic notation.
The fields are as follows:
image: image, A visual representation of the chess board showing the current piece arrangement.
board_state: string, A shortened FEN… See the full description on the dataset page: https://huggingface.co/datasets/bingbangboom/chess-puzzles-images-mini.puzzle-map
Puzzle-Map Dataset
The Puzzle-Map Dataset is a collection of images and annotations for research on computer vision applied to jigsaw puzzles.
Directory Structure
masks/
*.png | *.jpg
masks-raw/
*.png | *.jpg
pieces/
*.png | *.jpg
annotations.json
puzzles/
*.png | *.jpg
annotations.json
Description
masks/ – Processed masks used for data augmentation.
masks-raw/ – Original masks used to generate the processed masks.… See the full description on the dataset page: https://huggingface.co/datasets/pablo-moreira/puzzle-map.chess-puzzles-images-large
Dataset Card for Chess Puzzles Images (large)
This dataset contains 1,249,999 chess board positions in JPG format, derived from Lichess puzzles. Each image is accompanied by a shortened FEN string, indication for the color to play as, castling and en passant availability, and best moves in standard algebraic notation.
The fields are as follows:
image: image, A visual representation of the chess board showing the current piece arrangement.
board_state: string, A shortened FEN… See the full description on the dataset page: https://huggingface.co/datasets/bingbangboom/chess-puzzles-images-large.puzzles-for-vision-llmPuzzleVQAPuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Abstract Visual Patterns
rebus-puzzles
Rebus Dataset
The Rebus Dataset is a collection of 221 rebus puzzle images, each annotated with corresponding textual solutions and metadata.It was introduced as part of the paper:
Reasoning Riddles: How Explainability Reveals Cognitive Limits in Vision-Language ModelsPrahitha Movva, 2025arXiv:2510.02780
The dataset is designed to support research in visual reasoning, multimodal interpretability, and cognitive evaluation of vision–language models.
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/pmovva/rebus-puzzles.Jigsaw-Puzzles
Jigsaw-Puzzles Dataset
Jigsaw-Puzzles is a novel benchmark consisting of 1,100 carefully curated real-world images with high spatial complexity, designed to rigorously evaluate Vision-Language Models' (VLMs) spatial perception, structural understanding, and reasoning capabilities. The dataset minimizes reliance on domain-specific knowledge to better isolate and assess general spatial reasoning, positioning itself as a challenging and diagnostic benchmark for advancing spatial… See the full description on the dataset page: https://huggingface.co/datasets/zesen01/Jigsaw-Puzzles.Puzzle_Perception
Puzzle Perception — Segmentation + pVQA
A single table over two tasks on synthetic puzzle images:
Segmentation — per-pixel masks over chess, maze and tower-of-hanoi under one
unified 30-class label space.
pVQA — multiple-choice perception probes over chess and N-Queens boards.
Every row carries the same 13 columns; the type column says which task it
belongs to, and columns that do not apply are null.
from datasets import load_dataset
ds =… See the full description on the dataset page: https://huggingface.co/datasets/PuzzleComm/Puzzle_Perception.Sodoku_Puzzle_GeneratorDeveloping an MLP-Based AI/ML Model for Sudoku Puzzle Solving
Introduction to AI/ML Sudoku Solvers
Sudoku, a widely recognized logic-based combinatorial number-placement puzzle, presents a compelling challenge for Artificial Intelligence and Machine Learning models. The objective of Sudoku is to populate a 9x9 grid, which is further subdivided into nine 3x3 subgrids, with digits ranging from 1 to 9. The fundamental constraint is that each digit must appear exactly once within each row, each… See the full description on the dataset page: https://huggingface.co/datasets/MartialTerran/Sodoku_Puzzle_Generator.puzzle_manipulation_datasets
Puzzle Manipulation Datasets
This dataset is a collection of logged data from the teleoperation of ergoCub manipulating a puzzle.
Dataset Details
Dataset Description
This dataset contains data from the teleoperation of ergoCub collected in an experiment where the robot first unsolves then solves a puzzle.
Curated by: Giovanni Fregonese (@giotherobot)
License: CC BY 4.0
Dataset Creation
Curation Rationale
The dataset was created… See the full description on the dataset page: https://huggingface.co/datasets/ami-iit/puzzle_manipulation_datasets.translated_visual_puzzles_with_questionimage-with-puzzleLEGO-Puzzleschess-puzzle-tasks-visionpuzzlevqatranslated_visual_puzzlespuzzlevqa_easyr1_hard_smaller_allDeepseek-zebra-puzzlePuzzleVQApuzzlevqa_easyr1_hard_small_allLEGO-Puzzlespuzzlevqa_easyr1_hardpuzzlevqa_easyr1_easy_200puzzlevqa_easyr1PuzzleVQA-TRpuzzlevqa_easyr1_easy_200_small
