saad1926q/8-puzzle
8-puzzle / 3x3 sliding puzzle Fixed 3x3 / 8-puzzle evaluation data and teacher-rollout datasets for sliding-puzzle reasoning experiments. Configs rl: 35 unique run-1 training puzzles at optimal depths 2-6, excluding all 31 fixed evaluation boards. It uses the run_1 split and contains only board and optimal_length. eval: 31 fixed evaluation puzzles, with exactly one puzzle at every optimal distance from 1 through 31. sft-source: 200 fresh boards, balanced with 20… See the full description on the dataset page: https://huggingface.co/datasets/saad1926q/8-puzzle.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face