CoolFace
20 results

stepfun

stepfun-ai /Step-3.5-Flash-SFT Step-3.5-Flash-SFT Step-3.5-Flash-SFT is a general-domain supervised fine-tuning release for chat models. This repository keeps the full training interface in one place: json/: canonical raw training data tokenizers/: tokenizer snapshots for Step-3.5-Flash and Qwen3, released to preserve chat-template alignment compiled/: tokenizer-specific compiled shards for StepTronOSS training Data Format Each raw shard is a JSON file whose top level is a list of examples.… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/Step-3.5-Flash-SFT.text-generation1M<n<10M347 likes6.3k downloads6mo agoHugging Facestepfun-ai /GEdit-BenchDataset for Step1X-Edit: A Practical Framework for General Image Editing. This dataset is a new benchmark, grounded in real-world usages is developed to support more authentic and comprehensive evaluation of image editing models. Code imageimage-to-image1K<n<10K32 likes2.5k downloads1y agoHugging Facestepfun-ai /PaCoRe-Train-8k PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning Read the Paper | GitHub Repository | Download Models | Training Data 📖 Overview We introduce PaCoRe (Parallel Coordinated Reasoning), a framework that shifts the driver of inference from sequential depth to coordinated parallel breadth, breaking the model context limitation and massively scaling test time compute: Think in Parallel: PaCoRe launches massive parallel exploration… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/PaCoRe-Train-8k.texttext-generation1K<n<10K80 likes720 downloads8mo agoHugging Facestepfun-ai /Step-Video-T2V-EvalThis dataset contains the data of the paper Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model. Code: https://github.com/stepfun-ai/Step-Video-T2V Project page: https://yuewen.cn/videos videotext-to-videon<1K1 likes296 downloads2y agoHugging Facestepfun-ai /StepEval-Audio-Paralinguistic StepEval-Audio-Paralinguistic Dataset Paper: Step-Audio 2 Technical ReportCode: https://github.com/stepfun-ai/Step-Audio2Project Page: https://www.stepfun.com/docs/en/step-audio2 Overview StepEval-Audio-Paralinguistic is a speech-to-speech benchmark designed to evaluate AI models' understanding of paralinguistic information in speech across 11 distinct dimensions. The dataset contains 550 carefully curated and annotated speech samples for assessing capabilities beyond… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/StepEval-Audio-Paralinguistic.audion<1K12 likes276 downloads1y agoHugging Facestepfun-ai /StepEval-Audio-Toolcall StepEval-Audio-Toolcall Paper: Step-Audio 2 Technical ReportCode: https://github.com/stepfun-ai/Step-Audio2Project Page: https://www.stepfun.com/docs/en/step-audio2 Dataset Description StepEval Audio Toolcall evaluates the invocation performance of four tool types. For each tool, the benchmark contains approximately 200 multi-turn dialogue sets for both positive and negative scenarios: Positive samples: The assistant is required to invoke the specified tool in the… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/StepEval-Audio-Toolcall.audio-text-to-text7 likes267 downloads1y agoHugging Face