stepfun
Datasets
All datasets matching “stepfun”Step-3.5-Flash-SFT
Step-3.5-Flash-SFT
Step-3.5-Flash-SFT is a general-domain supervised fine-tuning release for chat models.
This repository keeps the full training interface in one place:
json/: canonical raw training data
tokenizers/: tokenizer snapshots for Step-3.5-Flash and Qwen3, released to preserve chat-template alignment
compiled/: tokenizer-specific compiled shards for StepTronOSS training
Data Format
Each raw shard is a JSON file whose top level is a list of examples.… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/Step-3.5-Flash-SFT.GEdit-BenchDataset for Step1X-Edit: A Practical Framework for General Image Editing.
This dataset is a new benchmark, grounded in real-world usages is developed to support more authentic and comprehensive evaluation of image editing models.
Code
PaCoRe-Train-8k
PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning
Read the Paper | GitHub Repository | Download Models | Training Data
📖 Overview
We introduce PaCoRe (Parallel Coordinated Reasoning), a framework that shifts the driver of inference from sequential depth to coordinated parallel breadth, breaking the model context limitation and massively scaling test time compute:
Think in Parallel: PaCoRe launches massive parallel exploration… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/PaCoRe-Train-8k.Step-Video-T2V-EvalThis dataset contains the data of the paper Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model.
Code: https://github.com/stepfun-ai/Step-Video-T2V
Project page: https://yuewen.cn/videos
StepEval-Audio-Paralinguistic
StepEval-Audio-Paralinguistic Dataset
Paper: Step-Audio 2 Technical ReportCode: https://github.com/stepfun-ai/Step-Audio2Project Page: https://www.stepfun.com/docs/en/step-audio2
Overview
StepEval-Audio-Paralinguistic is a speech-to-speech benchmark designed to evaluate AI models' understanding of paralinguistic information in speech across 11 distinct dimensions. The dataset contains 550 carefully curated and annotated speech samples for assessing capabilities beyond… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/StepEval-Audio-Paralinguistic.StepEval-Audio-Toolcall
StepEval-Audio-Toolcall
Paper: Step-Audio 2 Technical ReportCode: https://github.com/stepfun-ai/Step-Audio2Project Page: https://www.stepfun.com/docs/en/step-audio2
Dataset Description
StepEval Audio Toolcall evaluates the invocation performance of four tool types. For each tool, the benchmark contains approximately 200 multi-turn dialogue sets for both positive and negative scenarios:
Positive samples: The assistant is required to invoke the specified tool in the… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/StepEval-Audio-Toolcall.
