datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
thinking-cap-tier-raw-traces
Thinking Cap Tier Raw Traces (TCS v4)
[!IMPORTANT]
Dataset Release v1.2 (Sept 2026) — Clean Delimiters & Zero-Padding Architecture:
All 38,158 candidate reasoning traces across all 4 tiers (candidates_low.jsonl, candidates_mid.jsonl, candidates_high.jsonl, candidates_xhigh.jsonl) are 100% sanitized:
Zero batch-padding residues (<|pad|>): Completely purged across all records.
Strict Delimiter Integrity: Generation blocks cleanly separate thought deliberation tags… See the full description on the dataset page: https://huggingface.co/datasets/Davd-b01/thinking-cap-tier-raw-traces.DCAgent2_terminal_bench_2_laion_exp_tas_full_thinking_traces_20260102_04565520260728-001051-qwen3-30b-a3b-thinking-2507-qwen-opencode-v2-a708-tracesDCAgent2_terminal_bench_2_laion_exp_tas_interleaved_thinking_on_traces_20260102_050258qwen35-a3b-thinking-traces
Qwen3.5-35B-A3B Thinking Traces — SAE Training Data
Per-sentence L17 residual activations from Qwen/Qwen3.5-35B-A3B generating CoT on MMLU-Pro.
Stats
Model: Qwen/Qwen3.5-35B-A3B
Layer: L17 residual (~42% depth of 40-layer hybrid MoE)
Prompts: 2000 from MMLU-Pro test
Sentences: 41285
d_model: 2048
Activation dtype: float16
Purpose
Replication of Venhoff et al. 2025 (arXiv:2510.07364) "Base Models Know How to Reason, Thinking Models Learn When" applied to… See the full description on the dataset page: https://huggingface.co/datasets/caiovicentino1/qwen35-a3b-thinking-traces.claude-4-5-sonnet-thinking-stackexchange-overflow-32ep-32k-tracesexp_tas_interleaved_thinking_on_traces20260728-001051-qwen3-30b-a3b-thinking-2507-qwen-opencode-swebench-f9ff-traces20260728-001051-qwen3-30b-a3b-thinking-2507-qwen-opencode-tb2-11d7-tracesexp_tas_full_thinking_tracesthinking-traces-sft-100k
Thinking Traces SFT (100K)
100,000 ShareGPT-format conversations where the assistant shows explicit extended reasoning in <thinking> tags before giving a clean, structured final answer. Designed for training R1/o1-style reasoning models that separate the internal scratchpad from the public response.
Motivation
Standard SFT datasets train models to output correct answers. This dataset trains models to reason correctly — showing the full deliberation process before… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/thinking-traces-sft-100k.staqc-sandboxes-traces-terminus-2_Qwen3-4B-Thinking-2507DCAgent_dev_set_v2_laion_exp_tas_full_thinking_traceshard-coded-olmo-qwen3-vl-32b-thinking-traces-hand-filteredhard-coded-olmo-qwen3-vl-32b-thinking-tracesqwen3-14b-discard-thinking-tracesDCAgent2_bfcl-parity_laion_exp_tas_interleaved_thinking_on_traces_20260226_005613bfcl_parity_g1_min_episodes_e1_gpt_long_thinking_tacc_Qwen3_32B_20260417_194404-tracesdev_set_v2_claude_4_5_sonnet_thinking_stackexchange_overflow_32ep_32k_traces_201ae4588bDCAgent2_bfcl-parity_laion_exp_tas_full_thinking_traces_20260227_210016dev_set_v2_exp_tas_interleaved_thinking_on_traces_20260317_060628gsm8k-qwen3-235b-thinking-traces
gsm8k-qwen3-235b-thinking-traces
Teacher reasoning traces from Qwen3-235B-Thinking on 10 GSM8K problems for RST experiments
Dataset Info
Rows: 10
Columns: 10
Columns
Column
Type
Description
question
Value('string')
No description provided
answer
Value('string')
No description provided
metadata
Value('string')
No description provided
task_source
Value('string')
No description provided
id
Value('string')
No description provided… See the full description on the dataset page: https://huggingface.co/datasets/reasoning-degeneration-dev/gsm8k-qwen3-235b-thinking-traces.qwen3-14b-discard-thinking-elec-tracesDCAgent_dev_set_71_tasks_laion_exp_tas_interleaved_thinking_on_traces_20260102_073726DCAgent_dev_set_71_tasks_laion_exp_tas_full_thinking_traces_20260102_073856qwen3-14b-discard-thinking-games-tracesCodeRM-GRPO-2B-thinking-step700-test-traces
