puneetUMD/DSB-IFEval
DuplexSpeechBench–IFEval (DSB-IFEval) Evaluating implicit instruction following in full-duplex voice agents. ⚠️ Preprint — under review. Please cite it as a preprint (see below). DSB-IFEval tests whether a real-time voice agent can infer the turn-taking behavior a role implies — and execute it at the right moment on the conversational floor. It contains 1,038 evaluation cases built from 240 fixed user-side spoken interactions (8 assistant roles × 6 conversational probes × 5… See the full description on the dataset page: https://huggingface.co/datasets/puneetUMD/DSB-IFEval.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face