datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
riddlesense_plusplus
RiddleSense ++: Evaluating LLMs' Riddling abilities
Some cleaning and other modifications to make the riddle_sense dataset more suitable for training language models to generate riddles via standard text generation
Notable changes
reformatting to use special tags indicating the question and answer components
normalization: whitespace normalization, etc with clean-text
spell correction the original dataset has numerous spelling errors; these are fixed using BertChecker… See the full description on the dataset page: https://huggingface.co/datasets/pszemraj/riddlesense_plusplus.riddle_sense_standardizedriddlesenseplusplushungarian-riddles-benchmark
Hungarian Riddles Benchmark
Overview
This dataset is a cultural and reasoning benchmark based on 100 metaphorical, trivia-style Hungarian riddles.
The riddles are intentionally tricky and culturally grounded. They are designed to test answer correctness and reasoning quality, not only surface-level language fluency.
Dataset structure
Each row contains one riddle with reference material for evaluation.
Fields
ID – unique identifier
topic – general… See the full description on the dataset page: https://huggingface.co/datasets/boczkakaroly/hungarian-riddles-benchmark.
