datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
deepcoder-train-deepcoder-qwen4b-instruct-cont-temp0_6-32k-hsrun_step230-codeonly_truncationqwen4-exp-tiny-fidelity-root-v1
qwen4_exp random CPU fixture root
A root fidelity dataset in hidden form, produced by engines/tools/hf_capture.py from malaiwah/qwen4-exp-tiny-random-bf16.
The cut
the final hidden state handed to lm_head -- after the text model's final norm and immediately before the head matmul -- captured as the head module's input via torch.nn.Module.register_forward_pre_hook; replay applies the head ONLY (no final norm at replay time: the capture already sits after it). Same… See the full description on the dataset page: https://huggingface.co/datasets/malaiwah/qwen4-exp-tiny-fidelity-root-v1.qwen4-exp-tiny-cpu-repro-v1
Qwen4-Exp tiny CPU reproduction receipts
This is an artifact/receipt bundle, not training data and not a single root-format QFS dataset. All four readable synthetic documents are embedded in panel/panel.receipt.json.
Provenance and limitations
This is an independently generated, untrained random checkpoint inspired by Qwen/Qwen3.8-Flash-Next@de4b8e4d43b917e7706784d8bb445c9af86a3540, not a quantization, distillation, behavioral replica, or fine-tune. No source… See the full description on the dataset page: https://huggingface.co/datasets/malaiwah/qwen4-exp-tiny-cpu-repro-v1.Qwen4000-5000qwen_4b_base_deepsr_trainraconte-moi-10k-qwen4bqwen_4b_science_mix_train_gptossqwen_4b_medical_medmcqa_trainqwen4b-thinking-omni-l1_4-dpo-pairs-neg-othersqwen_4b_medical_medmcqa_train_40k_instruct_verifylawma-reasoning-qwen4b-v0qwen4b_2507_iter1olmo-3-preference-mix-deltas_reasoning-yolo_scottmix-chosen_qwen32b_rejected_qwen4b-DECONqwen_4b_science_base_train_gptossdm_qwen4b_datahpo_qwen4b_datadeepcoder_solver_genagg_qwen4b_instruct_50pqwen_4b_medical_medmcqa_train_10kint_stage1_qwen4b_instructqwen_4B_switch_explanationolmo-3-preference-mix-deltas_reasoning-chosen_qwen4b-yolo_scottmix-DECONmixed-trainabs-qwen4b-sft2e-6-samp16-all-flat-respgen-validQwen4B-MegaMath-pro-max-4096-len-verifierdan-caption-synth-qwen4bQwen4b-Judge-partial-qwen4b-no-thinking-omni-l5-step-scoreQwen8b-Judge-partial-qwen4b-no-thinking-omni-l5-step-scoreNemotron-Personas-USA-synthetic-records-10files-qa-vllm-qwen4b-instruct-2507-clarqadeepcoder-test-deepcoder-qwen4b-instruct-cont-temp0_6-32k-hsrun_step230-codeonly_truncationmixed-trainabs-qwen4b-sft2e-6-samp16-all-flat-respgenqwen_4b_medical_medmcqa_train_5k
