dev-sim
Datasets
All datasets matching “dev-sim”Simple-MathSteps-90K
Introducing Simple-MathSteps-90K:
An open source dataset of 93,325 elementary math problems with step-by-step solutions and multiple choice answers. Designed to enhance mathematical reasoning in models ranging from 1B to 13B parameters.
Key Features
93,325 Math Problems: Generated by paraphrasing the AQuA-RAT dataset using Qwen3 4B Instruct 2507, with a focus on consistency and quality.
Detailed Step-by-Step Solutions: Clear reasoning that breaks down problems… See the full description on the dataset page: https://huggingface.co/datasets/Raymond-dev-546730/Simple-MathSteps-90K.turnbench-dev-no-backchannel
TurnBench Dev - Backchannels Removed
A derivative of mundo-ai/turn-benchmark-dev
with every majority-annotated backchannel removed from the audio: 1853 backchannels
across 38 conversations, 2077.0 seconds in total, cut out of the
speaker's own channel and replaced by background noise taken from elsewhere in that same channel.
Everything else is the original recording, sample for sample. Same conversations, same duration,
same timeline, same annotator tracks, same speech -- only… See the full description on the dataset page: https://huggingface.co/datasets/JSALT2026-Conv-AI-Simulator/turnbench-dev-no-backchannel.lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-Shuffledlemonilia_LimaRP-Simple-CustomShareGPT-flatguard-splitlemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite
lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite
You should mask everything except the last turn. The only part that matters to teach the model is the last turn, as you are teaching it to always output thinking, no matter what the user feeds it.
It's setup to be trained like R1:
dedup_ablation_sim_threshold_0
