oreva/bass_music_benchmark
BASS: Benchmarking Audio LMs for Musical Structure and Semantic Reasoning BASS is a benchmark for evaluating music understanding and reasoning in audio language models. It comprises 2,658 questions across 12 tasks and 4 categories, covering 1,993 unique songs and over 138 hours of music. 🚀 Usage from datasets import load_dataset ds = load_dataset("oreva/bass_music_benchmark", "lyrics_transcription") ds = load_dataset("oreva/bass_music_benchmark"… See the full description on the dataset page: https://huggingface.co/datasets/oreva/bass_music_benchmark.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face