CoolFace
Datasetpublic

MostLime/chess-elite-uci

chess-elite-uci A transformer-ready dataset of ~7.8 million elite chess games, pre-tokenized in UCI notation with a deterministic 1977-token vocabulary. Built for training chess language models directly with no preprocessing required. Dataset Summary Field Value Total games 7,805,503 Average sequence length 94.24 tokens Max sequence length 255 tokens Vocabulary size 1,977 tokens Mean combined Elo 5,211 (~2,606 per player)… See the full description on the dataset page: https://huggingface.co/datasets/MostLime/chess-elite-uci.

sourceHugging Facecc-by-4.0updated 7mo agoView on Hugging Face
1likes42downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
MostLime/chess-elite-uci · CoolFace