CoolFace
Datasetpublic

AYipppp/gtow-llama-sft-v3

GTO Wizard — Heads-Up NL Hold'em 200BB — SFT dataset (v3) Supervised fine-tuning data for heads-up No-Limit Texas Hold'em, 200 big blinds deep. Each row is a single decision point: a natural-language description of the game state, paired with the game-theory-optimal action GTO Wizard chose in that spot. Intended for instruction-tuning a chat LLM to play HU 200BB poker (see the pokerbench agent it was built for). Schema Two flat columns: Column Description… See the full description on the dataset page: https://huggingface.co/datasets/AYipppp/gtow-llama-sft-v3.

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes62downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
AYipppp/gtow-llama-sft-v3 · CoolFace