Avesed/Qwen3.6-27B-DSpark-data
Qwen3.6-27B DSpark training data (on-policy, clean) On-policy conversations generated by Avesed/Qwen3.6-27B-W4A16 on a sha256-verified checkpoint, used to train Avesed/Qwen3.6-27B-DSpark. Each file is its own dataset config (they use different id schemes — integer vs zh_* string — so the viewer must keep them separate rather than merge into one table). config / file convs lang prompt source pb_pool94k_clean ~86k en PerfectBlend-style instruction mix general_onpolicy… See the full description on the dataset page: https://huggingface.co/datasets/Avesed/Qwen3.6-27B-DSpark-data.
configs: explicit split/path form (force viewer reprocess to 3 configs)
Add configs to separate the 3 files (fix dataset viewer schema conflict)
Add Chinese on-policy data (100000 convs, Magpie-zh + BELLE)
Add general on-policy data (416557 convs, clean ckpt)
Upload README.md with huggingface_hub
Clean on-policy data (sha256-verified target)
Delete pb_pool94k.jsonl with huggingface_hub
Upload README.md with huggingface_hub
On-policy DSpark training data (94k, Qwen3.6-27B-W4A16 responses)
Upload README.md with huggingface_hub
initial commit
