datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SpeechEditBench
SpeechEditBench
SpeechEditBench is a bilingual multi-attribute benchmark for
instruction-guided speech editing. Each example provides source speech and a
natural-language editing instruction, and the benchmark evaluates whether a
system can apply the requested edit while preserving the expected lexical
content.
Paper: SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing
Code and evaluator: github.com/daxintan-cuhk/SpeechEditBench… See the full description on the dataset page: https://huggingface.co/datasets/DiscreteSpeech/SpeechEditBench.speech_edit_acoustic
SpeechEdit Acoustic Retrieval Dataset
This dataset is an MTEB-formatted Any-to-Any (AT2A) composed audio retrieval adaptation of the acoustic_editing subset of DiscreteSpeech/SpeechEditBench.
Each query combines an original/source speech recording with a natural-language editing instruction, and the corpus contains the corresponding edited target speech recordings.
Schema
queries: id (string), audio (source audio), and text (edit instruction)
corpus: id (string)… See the full description on the dataset page: https://huggingface.co/datasets/deep9539/speech_edit_acoustic.Speech-Editing-Test
