datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
open-perfectblend-qwen3-4b-nonthinking
Open PerfectBlend Qwen3-4B Non-Thinking
This dataset contains 1,349,812 training conversations derived from mlabonne/open-perfectblend. Each assistant turn was regenerated sequentially with Qwen/Qwen3-4B, conditioned on the preceding conversation, with thinking disabled.
Data preparation
Empty or otherwise invalid source conversations were removed before a deterministic train/evaluation split. The split used seed 42 and a held-out fraction of 0.05. Only the 1,349… See the full description on the dataset page: https://huggingface.co/datasets/linYD0718/open-perfectblend-qwen3-4b-nonthinking.ShareGPT-Qwen3-4B-T0.7-NonThinking-Regen
ShareGPT Qwen3-4B T0.7 Non-Thinking Regen
ShareGPT conversations regenerated with Qwen/Qwen3-4B in non-thinking mode.
Generation settings
temperature: 0.7
top-p: 0.8
top-k: 20
min-p: 0
max tokens: 4096
reasoning: disabled
system message: You are a helpful assistant.
The source contained 36,943 rows. Regeneration produced 36,936 successful rows,
skipped 7 rows, and recorded no generation errors. The dataset contains the
successful rows only.
Each JSONL record has… See the full description on the dataset page: https://huggingface.co/datasets/TY233/ShareGPT-Qwen3-4B-T0.7-NonThinking-Regen.
