CoolFace
Datasetpublic

jevonmao/gtow-llama-sft-v3

GTO Wizard — Heads-Up NL Hold'em 200BB — SFT dataset (v3) Supervised fine-tuning data for heads-up No-Limit Texas Hold'em, 200 big blinds deep. Each row is a single decision point: a natural-language description of the game state, paired with the game-theory-optimal action GTO Wizard chose in that spot. Intended for instruction-tuning a chat LLM to play HU 200BB poker (see the pokerbench agent it was built for). Schema Two flat columns: Column Description… See the full description on the dataset page: https://huggingface.co/datasets/jevonmao/gtow-llama-sft-v3.

sourceHugging Faceupdated 4mo agoView on Hugging Face
0likes76downloads
settings

This repository belongs to jevonmao on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namegtow-llama-sft-v3
visibilitypublic
licencenot set
gatedno
ownerjevonmao
Account settings
jevonmao/gtow-llama-sft-v3 · CoolFace