datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pixel-art-bench-v1
Pixel Art Benchmark Dataset (Source)
The Pixel Art Benchmark Dataset is a structured collection of pixel-art outputs generated by large language models (LLMs). Each sample consists of a discrete color palette and a grid-based representation of pixel art, along with generation metadata such as token usage, cost, and model provenance.
Each row in the dataset represents a single generated pixel-art sample.
Encoding Details
Each string in grid represents one row of pixels.… See the full description on the dataset page: https://huggingface.co/datasets/AINovice2005/pixel-art-bench-v1.pixel-art-bench-lite
🎨 Pixel Art Bench Lite
Pixel Art Bench Lite is a structured-output benchmark designed to evaluate small language models on their ability to generate valid, interpretable, and semantically meaningful JSON outputs under strict constraints.
The benchmark is based on Pixel Art Bench focuses on pixel art generation over a fixed 24×24 grid, requiring models to produce outputs that are syntactically correct but also visually coherent.
While many benchmarks evaluate free-form text… See the full description on the dataset page: https://huggingface.co/datasets/AINovice2005/pixel-art-bench-lite.
