datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
s1K-1.1-dataforge-testing-20251216-123019
Dataset Card for lewtun/s1K-1.1-dataforge-testing-20251216-123019
Dataset Summary
Synthetic data generated by DataForge:
Model: Qwen/Qwen3-4B-Instruct-2507 (main)
Source dataset: simplescaling/s1K-1.1 (train split).
Generation config: temperature=0.7, top_p=0.8, top_k=20, max_tokens=4096, model_max_context=32768
Speculative decoding: disabled
System prompt: None
User prompt: Column question
The run produced 1,000 samples and generated 3,406,836 (~3.4M) tokens.
You can… See the full description on the dataset page: https://huggingface.co/datasets/lewtun/s1K-1.1-dataforge-testing-20251216-123019.s1K-1.1-dataforge-testing-20251216-142704
Dataset Card for lewtun/s1K-1.1-dataforge-testing-20251216-142704
Dataset Summary
Synthetic data generated by DataForge:
Model: Qwen/Qwen3-4B-Instruct-2507 (main)
Source dataset: simplescaling/s1K-1.1 (train split).
Generation config: temperature=0.7, top_p=0.8, top_k=20, max_tokens=4096, model_max_context=32768
Speculative decoding: disabled
System prompt: None
User prompt: Column question
The run produced 10 samples and generated 30,174 tokens.
You can load the… See the full description on the dataset page: https://huggingface.co/datasets/lewtun/s1K-1.1-dataforge-testing-20251216-142704.
