datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Malaysian-Multiturn-Chat-Assistant
Malaysian-Multiturn-Chat-Assistant
Generate synthetic multi-turn chat assistant with complex system prompt using mesolitica/Malaysian-Qwen2.5-72B-Instruct.
After that generate synthetic voice using mesolitica/Malaysian-Dia-1.6B also verified with Force Alignment to make sure the pronunciations almost correct.
A conversation must at least have 2 audio. We follow chat template from Qwen/Qwen2-Audio-7B-Instruct.
how to prepare the dataset
huggingface-cli download \… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Malaysian-Multiturn-Chat-Assistant.synthetic-wakeword-voice_assistant
synthetic-wakeword-voice_assistant
Synthetic wake-word audio for training and benchmarking OVOS wake-word
plugins, covering the phrase "voice assistant".
Every clip is machine-generated: text-to-speech synthesis followed by voice
conversion to simulate multiple speakers. No human recording is included, and
no natural voice is reproduced. Machine-generated audio carries no copyright
of its own, so this dataset is published CC-BY-4.0 and is free to use,
redistribute and build on… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/synthetic-wakeword-voice_assistant.synthetic-wakeword-home_assistant
synthetic-wakeword-home_assistant
Synthetic wake-word audio for training and benchmarking OVOS wake-word
plugins, covering the phrase "home assistant".
Every clip is machine-generated: text-to-speech synthesis followed by voice
conversion to simulate multiple speakers. No human recording is included, and
no natural voice is reproduced. Machine-generated audio carries no copyright
of its own, so this dataset is published CC-BY-4.0 and is free to use,
redistribute and build on… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/synthetic-wakeword-home_assistant.assistant-bench
Assistant Bench
31-turn multi-turn speech-to-speech benchmark for evaluating voice AI models as a personal assistant handling flights, email, calendar, and reminders.
Part of Audio Arena, a suite of 6 benchmarks spanning 221 turns across different domains. Built by Arcada Labs.
Leaderboard | GitHub | All Benchmarks
Dataset Description
The model acts as a personal assistant managing flight bookings, email composition, calendar events, and reminders. Turns include dual… See the full description on the dataset page: https://huggingface.co/datasets/arcada-labs/assistant-bench.VoiceStudy-Assistant-audio
