CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01gpt-omni /VoiceAssistant-400Kaudio100K<n<1M100 likes4.1k downloads2y agoHugging Face02MathLLMs /VoiceAssistant-Eval 🔥 VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing [🌐 Homepage] [🔮 Visualization] [💻 Github] [📖 Paper] [📊 Leaderboard ] [📊 Detailed Leaderboard ] [📊 Roleplay Leaderboard ] 🚀 Data Usage from datasets import load_dataset for split in ['listening_general', 'listening_music', 'listening_sound', 'listening_speech', 'speaking_assistant', 'speaking_emotion', 'speaking_instruction_following'… See the full description on the dataset page: https://huggingface.co/datasets/MathLLMs/VoiceAssistant-Eval.textquestion-answering10K<n<100K12 likes653 downloads11mo agoHugging Face03TigreGotico /synthetic-wakeword-voice_assistant synthetic-wakeword-voice_assistant Synthetic wake-word audio for training and benchmarking OVOS wake-word plugins, covering the phrase "voice assistant". Every clip is machine-generated: text-to-speech synthesis followed by voice conversion to simulate multiple speakers. No human recording is included, and no natural voice is reproduced. Machine-generated audio carries no copyright of its own, so this dataset is published CC-BY-4.0 and is free to use, redistribute and build on… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/synthetic-wakeword-voice_assistant.audioaudio-classification1K<n<10K0 likes94 downloads20d agoHugging Face04Muvels /VoiceAssistant-400K-mimi VoiceAssistant-400K with Mimi Tokens This dataset is a processed version of gpt-omni/VoiceAssistant-400K with audio codec conversions. Processing Each sample has been processed to add: answer_audio: Decoded audio waveform from SNAC tokens (24kHz Audio feature) answer_mimi: Re-encoded audio using Kyutai's Mimi codec (32 codebooks) Columns Column Type Description split_name string Original split name index int Sample index round int Conversation… See the full description on the dataset page: https://huggingface.co/datasets/Muvels/VoiceAssistant-400K-mimi.audiotext-to-speech1K<n<10K0 likes47 downloads10mo agoHugging Face05SamSoko83 /VoiceAssistant-Eval 🔥 VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing [🌐 Homepage] [🔮 Visualization] [💻 Github] [📖 Paper] [📊 Leaderboard ] [📊 Detailed Leaderboard ] [📊 Roleplay Leaderboard ] 🚀 Data Usage from datasets import load_dataset for split in ['listening_general', 'listening_music', 'listening_sound', 'listening_speech', 'speaking_assistant', 'speaking_emotion', 'speaking_instruction_following'… See the full description on the dataset page: https://huggingface.co/datasets/SamSoko83/VoiceAssistant-Eval.textquestion-answering10K<n<100K0 likes34 downloads3mo agoHugging Face06yoleneyao /VoiceAssistant-400Kaudio100K<n<1M0 likes2 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.