datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dualturn-otospeech-turn-taking
OtoSpeech Turn-Taking
Official DualTurn release of the otospeech corpus, with per-frame turn-taking labels and
Mimi speech codec features. Each row is one full conversation. Frame rate 12.5 Hz (80 ms per frame).
Paper: DualTurn: Learning Turn-Taking from Dual-Channel Generative Speech Pretraining
Training code: github.com/anyreachai/dualturn
Model checkpoint: anyreach-ai/dualturn-qwen2.5-mimi-0.5B
Splits
Split
Sessions
train
896
val
111
test
113… See the full description on the dataset page: https://huggingface.co/datasets/anyreach-ai/dualturn-otospeech-turn-taking.dualturn-switchboard-turn-taking
Switchboard Turn-Taking
Official DualTurn release of the switchboard corpus, with per-frame turn-taking labels and
Mimi speech codec features. Each row is one full conversation. Frame rate 12.5 Hz (80 ms per frame).
Paper: DualTurn: Learning Turn-Taking from Dual-Channel Generative Speech Pretraining
Training code: github.com/anyreachai/dualturn
Model checkpoint: anyreach-ai/dualturn-qwen2.5-mimi-0.5B
Splits
Split
Sessions
train
1986
val
295
test
138… See the full description on the dataset page: https://huggingface.co/datasets/anyreach-ai/dualturn-switchboard-turn-taking.candor-turntaking-annotations
CANDOR - Turn-Taking Annotations
Speech transcription and turn-taking annotation dataset built from the CANDOR corpus using NVIDIA Canary-Qwen2.5B ASR.
Dataset Description
This dataset contains 172,591 transcribed speech segments from the CANDOR conversational speech corpus (1,656 conversations). Each segment is a per-speaker utterance with Canary ASR transcript, designed for turn-taking prediction research.
Source
Audio corpus: CANDOR (English conversational… See the full description on the dataset page: https://huggingface.co/datasets/hiraki/candor-turntaking-annotations.
