datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
LoGos-Rollout-1K
LoGos-Rollout-1K
Resources
Paper: Mixing Expert Knowledge: Bring Human Thoughts Back To the Game of Go
GitHub Repository: https://github.com/Entarochuan/LoGos
Associated Model: LoGos-7B
Citation
@misc{ma2026mixingexpertknowledgebring,
title={Mixing Expert Knowledge: Bring Human Thoughts Back To the Game of Go},
author={Yichuan Ma and Linyang Li and Yongkang Chen and Peiji Li and Jiasheng Ye and Qipeng Guo and Dahua Lin and Kai Chen}… See the full description on the dataset page: https://huggingface.co/datasets/YichuanMa/LoGos-Rollout-1K.LogosForge-scored-sft-v1
LogosForge-scored-sft-v1
This dataset is a scored Supervised Fine-Tuning (SFT) distillation corpus built on top of the Natural Reasoning question set.It is constructed in two stages:
First, a large-scale reasoning-oriented teacher model (gpt-oss-120B-high) is used to generate distilled student responses, including explicit chain-of-thought reasoning, for natural reasoning questions.
Second, these distilled responses are evaluated by a separate instruction-following model… See the full description on the dataset page: https://huggingface.co/datasets/Jackrong/LogosForge-scored-sft-v1.logos-corpus
Logos v1.0 Corpus
The training corpus for Logos — the first systems programming language built for AI, not borrowed from humans.
Logos v1.0 shipped 2026-05-21. This dataset is the curated corpus of .logos programs as of v1.0.0 — cookbook recipes, algorithm examples, standard library, and integration-test fixtures. Every program in this dataset:
Typechecks under logosc v1.0.0 (Hindley-Milner inference with effect inference + polymorphism)
Refinement contracts (where present) lower… See the full description on the dataset page: https://huggingface.co/datasets/byShammy/logos-corpus.
