datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qwen4-exp-tiny-fidelity-root-v1
qwen4_exp random CPU fixture root
A root fidelity dataset in hidden form, produced by engines/tools/hf_capture.py from malaiwah/qwen4-exp-tiny-random-bf16.
The cut
the final hidden state handed to lm_head -- after the text model's final norm and immediately before the head matmul -- captured as the head module's input via torch.nn.Module.register_forward_pre_hook; replay applies the head ONLY (no final norm at replay time: the capture already sits after it). Same… See the full description on the dataset page: https://huggingface.co/datasets/malaiwah/qwen4-exp-tiny-fidelity-root-v1.qwen4-exp-tiny-cpu-repro-v1
Qwen4-Exp tiny CPU reproduction receipts
This is an artifact/receipt bundle, not training data and not a single root-format QFS dataset. All four readable synthetic documents are embedded in panel/panel.receipt.json.
Provenance and limitations
This is an independently generated, untrained random checkpoint inspired by Qwen/Qwen3.8-Flash-Next@de4b8e4d43b917e7706784d8bb445c9af86a3540, not a quantization, distillation, behavioral replica, or fine-tune. No source… See the full description on the dataset page: https://huggingface.co/datasets/malaiwah/qwen4-exp-tiny-cpu-repro-v1.dm_qwen4b_datahpo_qwen4b_datamixed-trainabs-qwen4b-sft2e-6-samp16-all-flat-respgen-validQwen4b-Judge-partial-qwen4b-no-thinking-omni-l5-step-scoreqwen-4b-step0-dataQwen8b-Judge-partial-qwen4b-no-thinking-omni-l5-step-scoredeepcoder-train-qwen4b-instr-tok-mean_bs64_roll4-run3_step50-codeonly_truncation_inferencepromptQwen4b-2507-Judge-partial-qwen4b-instruct-2507-omni-l7-step-scoredeepscaler-easy_qwen4b-og-abstraction-sft-5epoch-0702Qwen4b-Instruct-Judge-partial-qwen4b-instruct-omni-l7-step-scoreslateral-think-rl-5hint-diff-qwen4b-verldeepscaler-easy_qwen4b-abstraction-sft-nosols-sftinit-0702_s4mixed-trainabs-qwen4b-sft1e-5-samp16-all-flat-respgen-validmixed-trainabs-qwen4b-sft5e-6-Qwen3-4B-AWQ-samp16-max100-validQwen14b-Judge-partial-qwen4b-no-thinking-omni-l5-step-scoredeepscaler-easy_qwen4b-abstraction-sft-nosols-0702_s7deepscaler-easy_qwen4b-abstraction-sft-nosols-0702_s0deepscaler-easy_qwen4b-abstraction-sft-nosols-0702_s2deepscaler-easy_qwen4b-abstraction-sft-nosols-sftinit-0702deepscaler-easy_qwen4b-abstraction-sft-nosols-0702deepscaler-hard_qwen4b-abstraction-sft-nosols-sftinit-0702_s0deepscaler-easy_qwen4b-abstraction-sft-nosols-0702_s3deepscaler-easy_qwen4b-og-abstraction-sft-5epoch-0702_s4deepscaler-easy_qwen4b-og-abstraction-sft-5epoch-0702_s3deepscaler-easy_qwen4b-og-abstraction-sft-5epoch-0702_s1deepscaler-easy_qwen4b-og-abstraction-sft-5epoch-0702_s7mixed-trainabs-qwen4b-sft5e-6-samp16-all-flat-respgen-validqwen-4b-noveltybench-comprehensive-evaluation
