CoolFace
Datasetpublic

riddickz/llada-sudoku-violation-pairs-50k-v1

LLaDA Sudoku Violation Pairs (50k v1) SFT pairs for training a diffusion language model (LLaDA-8B-Instruct) to diagnose constraint violations and emit a corrected solution in a single response — the "diagnose + correct in one shot" objective. Each row is a (prompt, response, meta) triple where: prompt is the LLaDA-native chat-template prefix containing the sudoku puzzle and the model's wrong answer (one full assistant turn already written), with a fresh assistant header opened… See the full description on the dataset page: https://huggingface.co/datasets/riddickz/llada-sudoku-violation-pairs-50k-v1.

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
0likes13downloads

riddickz/llada-sudoku-violation-pairs-50k-v1 · main · files are served by the source, never re-hosted here