CoolFace
Datasetpublic

saad1926q/8-puzzle

8-puzzle / 3x3 sliding puzzle Fixed 3x3 / 8-puzzle evaluation data and teacher-rollout datasets for sliding-puzzle reasoning experiments. Configs rl: 35 unique run-1 training puzzles at optimal depths 2-6, excluding all 31 fixed evaluation boards. It uses the run_1 split and contains only board and optimal_length. eval: 31 fixed evaluation puzzles, with exactly one puzzle at every optimal distance from 1 through 31. sft-source: 200 fresh boards, balanced with 20… See the full description on the dataset page: https://huggingface.co/datasets/saad1926q/8-puzzle.

sourceHugging Faceupdated 2d agoView on Hugging Face
0likes470downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
saad1926q/8-puzzle · CoolFace