datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
running-coach-sft
Running Coach SFT
Instruction-tuning data for a distance-running coaching assistant. Every pace,
split, and race-equivalent in the corpus is computed from a Daniels/Gilbert VDOT
implementation rather than written into a template, so the numbers are internally
consistent across all 1,500 examples.
Why this exists
Coaching corpora scraped from forums and blogs teach a model the register of
coaching without the arithmetic underneath it. A model that interpolates… See the full description on the dataset page: https://huggingface.co/datasets/hoodarunner/running-coach-sft.chess-coach-benchmark
Chess coach benchmark
A comprehensive, zero-leakage benchmark measuring the performance of local chess-coaching fine-tunes against frontier models.
What it is
This benchmark evaluates models on held-out chess positions, measuring their ability to provide tier-calibrated, engine-grounded chess coaching. It scores models based on deterministic objective metrics (move soundness, absence of engine jargon, and verification of board facts) alongside a blinded… See the full description on the dataset page: https://huggingface.co/datasets/khoilamalphaai/chess-coach-benchmark.chess-coach-move-review
Chess coach move-review SFT dataset
Supervised fine-tuning data for one specific, trained behavior: given a chess
position and the student's rating tier (Beginner, Intermediate, or Advanced),
select the tier-appropriate instructive move and tag it with a short principle,
for example "Nf3, develop toward the center."
That single move choice is the trained objective, and it is deterministically
checkable. The plain-English explanation rendered beside the move is a secondary… See the full description on the dataset page: https://huggingface.co/datasets/khoilamalphaai/chess-coach-move-review.chess-coach-grand-eval
Chess Coach — Grand Eval (comprehensive leaderboard)
One fresh, apples-to-apples comparison of every model in the chess move-review
coaching project — our tuned specialists, the untuned baselines, and the full frontier
lineup — on the same held-out validation slice (120 positions × 3 tiers
= 360 scenarios), scored with two independent layers:
Deterministic moat metrics (free, python-chess over pre-computed Stockfish/Maia
facts): tier-fit, distinct-moves-per-level… See the full description on the dataset page: https://huggingface.co/datasets/khoilamalphaai/chess-coach-grand-eval.chess-coach-turningpoints
Chess Coach – Turning Point Explanations Dataset
Overview
This repository contains a curated, engine-grounded dataset for training language models to explain chess mistakes and turning points in a human coaching style.
The goal is explainability and pedagogy, not move calculation or engine strength.
What this dataset is (and is not)
✅ This dataset is for
Training LLMs to explain evaluation swings
Teaching coaching tone, structure, and pedagogy… See the full description on the dataset page: https://huggingface.co/datasets/suman-kalavagunta/chess-coach-turningpoints.chess-coach-v6
Chess Coach v6 (deep-verified training labels)
The current data frontier for the chess-instructor-llm coach: a foundational,
data-first rebuild of the training LABELS (the move plus full provenance), deep-verified
with Stockfish 17 (a two-depth root search with agreement bands), Syzygy tablebases
(endgames of seven pieces or fewer), and Maia-2 human-likelihood. It feeds the
downstream preference (DPO) and engine-distillation retrains.
This dataset is NOT the shipped SFT set. The… See the full description on the dataset page: https://huggingface.co/datasets/khoilamalphaai/chess-coach-v6.garmin-mexican-fitness-coach-sft
🇲🇽 Garmin Mexican Fitness Coach SFT Dataset
Dataset sintético de alta fidelidad para el ajuste fino instruccional (Supervised Fine-Tuning / LoRA) de modelos de lenguaje pequeños (e.g., Llama 3.2 1B/3B, Qwen 2.5 1.5B/3B, SmolLM2), diseñado para dotarlos de la personalidad, modismos y tono enérgico de un Coach de Alto Rendimiento Mexicano con fundamento fisiológico estricto (Firstbeat Technologies & Garmin Connect).
📊 Resumen del Dataset
Tamaño total: 400… See the full description on the dataset page: https://huggingface.co/datasets/GerardoMayel/garmin-mexican-fitness-coach-sft.Coach-1.2k
