datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VoiceAssistant_unitsWe use mHuBERT and K-means Model to convert answer from VoiceAssistant (synthesized by vits) to units.
*pred* indicates the answer from the speech model
*gt* indicates the answer from the ground-truth
For usage, please refer to Open-Omni-Nexus
Voice_assistantmental_health_voice_assistantMental_health_VoiceAssistantOnboard-Vehicle-Voice-Assistant
