datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ATRI_Voice_DatasetsLu_Yin_IndexTTS2_Sugar_Apple_Fairy_Tale_segments
sseu-bench
Abstract
Recently, Large Audio Language Models (LALMs) have progressed rapidly, demonstrating their strong efficacy in universal audio understanding through cross-modal integration.
To evaluate LALMs' audio understanding performance, researchers have proposed different benchmarks.
However, key aspects for real-world interactions are underexplored in existing benchmarks, i.e., audio signals typically contain both speech and non-speech components, and energy levels of these… See the full description on the dataset page: https://huggingface.co/datasets/apple121/sseu-bench.Yanamianna_Voice_Dataset
该数据集为《败犬女角太多了》的小八语言集.
MMAU-Pro-CtrlMusicPro-7kMusicPro-7k Dataset:A professional, versatile, and high-quality Video-to-Music Dataset, specifically focused on film music. 🎵Proposed in the FilmComposer project, aimed at advancing research in music production & video-to-music generation. 🎶
Licenseplease fill the MusicPro-7k Terms of Use and send email to heqi389973717@gmail.com.
✨ Dataset OverviewMusicPro-7k featuring about 7,418 samples, each with film clip, high-quality music, visual description, music description, main melody, and… See the full description on the dataset page: https://huggingface.co/datasets/apple-jun/MusicPro-7k.applejackmlpapplejackjack
