xiaomi-mimo
SpeechMMLU
Dataset Card for SpeechMMLU
Dataset Description
SpeechMMLU is an evaluation dataset designed to assess the knowledge capabilities of speech-language models. It is built based on the MMLU dataset, with entries filtered by subject and length, resulting in a total of 8,549 entries across 34 subjects. We synthesized the questions and answers into speech using commercial TTS with diverse voices. This dataset can be used for evaluating knowledge capabilities of speech-language… See the full description on the dataset page: https://huggingface.co/datasets/XiaomiMiMo/SpeechMMLU.MiMo-Audio-Evalset
Dataset Card for MiMo-Audio-Evalset
Dataset Description
This repository is a collection of multiple audio datasets used for evaluation in the MiMo-Audio-Eval toolkit. It includes a variety of datasets for tasks such as automatic speech recognition (AISHELL1, LibriSpeech), text-to-speech (SeedTTS), audio understanding and reasoning (MMAU and MMSU), and more.
Included Datasets
The following datasets are included in the MiMo-Audio-Evalset:
AISHELL1
LibriSpeech… See the full description on the dataset page: https://huggingface.co/datasets/XiaomiMiMo/MiMo-Audio-Evalset.
