CoolFace
Datasetpublic

riddickz/llada-sudoku-violation-pairs-50k-v1

LLaDA Sudoku Violation Pairs (50k v1) SFT pairs for training a diffusion language model (LLaDA-8B-Instruct) to diagnose constraint violations and emit a corrected solution in a single response — the "diagnose + correct in one shot" objective. Each row is a (prompt, response, meta) triple where: prompt is the LLaDA-native chat-template prefix containing the sudoku puzzle and the model's wrong answer (one full assistant turn already written), with a fresh assistant header opened… See the full description on the dataset page: https://huggingface.co/datasets/riddickz/llada-sudoku-violation-pairs-50k-v1.

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
0likes13downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
riddickz/llada-sudoku-violation-pairs-50k-v1 · CoolFace