datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Qwen3.6-27B-DSpark-data
Qwen3.6-27B DSpark training data (on-policy, clean)
On-policy conversations generated by Avesed/Qwen3.6-27B-W4A16
on a sha256-verified checkpoint, used to train Avesed/Qwen3.6-27B-DSpark.
Each file is its own dataset config (they use different id schemes — integer vs zh_*
string — so the viewer must keep them separate rather than merge into one table).
config / file
convs
lang
prompt source
pb_pool94k_clean
~86k
en
PerfectBlend-style instruction mix
general_onpolicy… See the full description on the dataset page: https://huggingface.co/datasets/Avesed/Qwen3.6-27B-DSpark-data.dspark-regen-instruct2507-opb-10000dspark-regen-input-opb-10000
