kmoss/qwen3-1.7b-traces
Qwen3-1.7B reasoning traces On-policy chain-of-thought rollouts from Qwen/Qwen3-1.7B in thinking mode, collected to study long-context KV-cache residency constraints (trainable sparse attention, in the InfLLM-V2 / NOSA line). Used as adaptation data: training a sparse+local attention variant on the model's own output distribution avoids the alignment tax that continued pretraining on external corpora imposes on a post-trained model. Caveats Not filtered for… See the full description on the dataset page: https://huggingface.co/datasets/kmoss/qwen3-1.7b-traces.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face