datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VoiceAssistant-400KVoiceAssistant-Eval
🔥 VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
[🌐 Homepage]
[🔮 Visualization]
[💻 Github]
[📖 Paper]
[📊 Leaderboard ]
[📊 Detailed Leaderboard ]
[📊 Roleplay Leaderboard ]
🚀 Data Usage
from datasets import load_dataset
for split in ['listening_general', 'listening_music', 'listening_sound', 'listening_speech',
'speaking_assistant', 'speaking_emotion', 'speaking_instruction_following'… See the full description on the dataset page: https://huggingface.co/datasets/MathLLMs/VoiceAssistant-Eval.synthetic-wakeword-voice_assistant
synthetic-wakeword-voice_assistant
Synthetic wake-word audio for training and benchmarking OVOS wake-word
plugins, covering the phrase "voice assistant".
Every clip is machine-generated: text-to-speech synthesis followed by voice
conversion to simulate multiple speakers. No human recording is included, and
no natural voice is reproduced. Machine-generated audio carries no copyright
of its own, so this dataset is published CC-BY-4.0 and is free to use,
redistribute and build on… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/synthetic-wakeword-voice_assistant.VoiceAssistant-400K-mimi
VoiceAssistant-400K with Mimi Tokens
This dataset is a processed version of gpt-omni/VoiceAssistant-400K with audio codec conversions.
Processing
Each sample has been processed to add:
answer_audio: Decoded audio waveform from SNAC tokens (24kHz Audio feature)
answer_mimi: Re-encoded audio using Kyutai's Mimi codec (32 codebooks)
Columns
Column
Type
Description
split_name
string
Original split name
index
int
Sample index
round
int
Conversation… See the full description on the dataset page: https://huggingface.co/datasets/Muvels/VoiceAssistant-400K-mimi.VoiceAssistant-Eval
🔥 VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
[🌐 Homepage]
[🔮 Visualization]
[💻 Github]
[📖 Paper]
[📊 Leaderboard ]
[📊 Detailed Leaderboard ]
[📊 Roleplay Leaderboard ]
🚀 Data Usage
from datasets import load_dataset
for split in ['listening_general', 'listening_music', 'listening_sound', 'listening_speech',
'speaking_assistant', 'speaking_emotion', 'speaking_instruction_following'… See the full description on the dataset page: https://huggingface.co/datasets/SamSoko83/VoiceAssistant-Eval.VoiceAssistant-400K
