datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VoiceAssistant-400K-SLAM-Omni
VoiceAssistant-400K (Modified)
This dataset is prepared for the reproduction of SLAM-Omni.
This is a single-round English spoken dialogue training dataset. For code and usage examples, please refer to the related GitHub repository: X-LANCE/SLAM-LLM (examples/s2s)
🔧 Modifications
Data Filtering: We removed samples with excessively long data.
Speech Response Tokens: We used CosyVoice to synthesize corresponding semantic speech tokens for the speech response. These… See the full description on the dataset page: https://huggingface.co/datasets/worstchan/VoiceAssistant-400K-SLAM-Omni.VoiceAssistant-400K-SLAM-Omni
VoiceAssistant-400K (Modified)
This dataset is prepared for the reproduction of SLAM-Omni.
This is a single-round English spoken dialogue training dataset. For code and usage examples, please refer to the related GitHub repository: X-LANCE/SLAM-LLM (examples/s2s)
🔧 Modifications
Data Filtering: We removed samples with excessively long data.
Speech Response Tokens: We used CosyVoice to synthesize corresponding semantic speech tokens for the speech response. These… See the full description on the dataset page: https://huggingface.co/datasets/mwei/VoiceAssistant-400K-SLAM-Omni.home-assistant-local-llm-voice-benchmark
Home Assistant Local LLM Voice Benchmark
Per-model tool-call accuracy and component latency for running a Home Assistant voice assistant against local LLMs.
Measured per-model tool-call accuracy and component latency for running a Home Assistant voice assistant against local LLMs.
Broken out by pipeline component rather than reported as one opaque round trip, so you can tell whether your latency is wake-word, speech-to-text, the model, or text-to-speech before you go optimizing… See the full description on the dataset page: https://huggingface.co/datasets/iBlessi/home-assistant-local-llm-voice-benchmark.
