CoolFace
20 results

rukh

chorcat /rukh-games-elite chorcat/rukh-games-elite Games from the Lichess Elite Database (2500+ against 2300+, no bullet), converted to legal UCI with the same schema as rukh-games-1800. Used for supervised fine-tuning on strong play. Part of Rukh, a chess language model built from scratch as a course on generative and agentic AI. Every derived dataset ships with the exact filters and counts of its manifest.json, so it can be regenerated with rukh data elite. Files File Bytes SHA-256… See the full description on the dataset page: https://huggingface.co/datasets/chorcat/rukh-games-elite.tabulartext-generation10M<n<100M0 likes59 downloads5d agoHugging Facechorcat /rukh-pairs-dpo chorcat/rukh-pairs-dpo Preference pairs from the multi-PV Stockfish lines of rukh-positions-eval: the best first move against a legal move at least 100 centipawns worse for the side to move (mates count as 10000). One pair per position, balanced by phase. Part of Rukh, a chess language model built from scratch as a course on generative and agentic AI. Every derived dataset ships with the exact filters and counts of its manifest.json, so it can be regenerated with rukh data… See the full description on the dataset page: https://huggingface.co/datasets/chorcat/rukh-pairs-dpo.tabulartext-generation10K<n<100K0 likes58 downloads5d agoHugging Facechorcat /rukh-puzzles-split chorcat/rukh-puzzles-split Lichess puzzles with rating deviation <= 100 and at least 100 plays, banded by difficulty (1000-1500, 1500-2000, 2000+) and split into test and train by a seeded hash of the puzzle id, each with the moves of the game it came from, for tactical evaluation and fine-tuning. Part of Rukh, a chess language model built from scratch as a course on generative and agentic AI. Every derived dataset ships with the exact filters and counts of its manifest.json, so… See the full description on the dataset page: https://huggingface.co/datasets/chorcat/rukh-puzzles-split.tabulartext-generation100K<n<1M0 likes50 downloads5d agoHugging Facechorcat /rukh-elo-bins chorcat/rukh-elo-bins A balanced sample of rukh-games-1800: up to n_per_bin games per 100-Elo bin of the average rating of both players, for Elo conditioning experiments. Part of Rukh, a chess language model built from scratch as a course on generative and agentic AI. Every derived dataset ships with the exact filters and counts of its manifest.json, so it can be regenerated with rukh data elo-bins. Files File Bytes SHA-256 games.parquet 65955015… See the full description on the dataset page: https://huggingface.co/datasets/chorcat/rukh-elo-bins.tabulartext-generation100K<n<1M0 likes47 downloads5d agoHugging Facechorcat /rukh-games-1800 chorcat/rukh-games-1800 Rated standard Lichess games with both players at 1800+ Elo, base time of at least 180 seconds, normal or time-forfeit terminations, 20 to 300 plies, converted from SAN to legal UCI. Partitioned by month: train on 2025-01, validate on 2025-02. Part of Rukh, a chess language model built from scratch as a course on generative and agentic AI. Every derived dataset ships with the exact filters and counts of its manifest.json, so it can be regenerated with… See the full description on the dataset page: https://huggingface.co/datasets/chorcat/rukh-games-1800.tabulartext-generation1M<n<10M0 likes47 downloads5d agoHugging Facechorcat /rukh-pairs-onpolicy chorcat/rukh-pairs-onpolicy Preference pairs built from the moves the model itself proposes. The positions are exactly those of rukh-pairs-dpo, so the only thing that differs between the two datasets is where the two moves came from -- which is what makes a DPO run on each of them a comparison. Four moves sampled per position at temperature 1.0, scored by Stockfish at a fixed depth of 10, keeping the best and the worst when they are at least 100 centipawns apart. Of the 13 838… See the full description on the dataset page: https://huggingface.co/datasets/chorcat/rukh-pairs-onpolicy.tabulartext-generation1K<n<10K0 likes47 downloads5d agoHugging Face