datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
neural-mathrock
Neural Math Rock Multimodal Emotion Dataset
Dataset Description
This corpus is a large-scale multimodal emotion classification dataset specifically developed for Music Information Retrieval (MIR) and emotional computational analysis within complex musical genres, predominantly Math Rock and Midwest Emo. The dataset consists of exactly 4,000 distinct full-length tracks structured directly from the validated metadata registry.
The primary objective of this corpus is… See the full description on the dataset page: https://huggingface.co/datasets/anggars/neural-mathrock.VoiceAssistant-Eval
🔥 VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
[🌐 Homepage]
[🔮 Visualization]
[💻 Github]
[📖 Paper]
[📊 Leaderboard ]
[📊 Detailed Leaderboard ]
[📊 Roleplay Leaderboard ]
🚀 Data Usage
from datasets import load_dataset
for split in ['listening_general', 'listening_music', 'listening_sound', 'listening_speech',
'speaking_assistant', 'speaking_emotion', 'speaking_instruction_following'… See the full description on the dataset page: https://huggingface.co/datasets/MathLLMs/VoiceAssistant-Eval.MoulSot-Full
MoulSot-Full Dataset
Dataset Summary
MoulSot-Full is a large-scale Moroccan Darija speech dataset containing in total 1,500 hours of speech audio. From this extensive corpus, a high-quality subset of approximately 80 hours has been carefully curated and transcribed. It was built entirely from publicly available YouTube content across 51 diverse channels (including vlogs, podcasts, interviews, and commentary) to capture real-world Moroccan Darija, including natural… See the full description on the dataset page: https://huggingface.co/datasets/MathematicianNLPer/MoulSot-Full.khanacademy-turkish-mathhard_math_wavHindi50_3french-math-asr-benchmarkasdhuawei-mathmatheusaleixomathbridge-audio
Dataset Card for MathBridge_Audio
Dataset Description
This dataset is an expansion of the MathBridge dataset by Kyudan. The original dataset gives translations from English text to LaTeX formatting with context before and after. Our dataset is a subset of 1,000 rows from MathBridge with added audio files of speakers reading the full text.
Curated by: Abigail Pitcairn
Funded by: NSF and University of Southern Maine AIIR Lab
Language: English (en)
License: open-source… See the full description on the dataset page: https://huggingface.co/datasets/abby1492/mathbridge-audio.MathSpeechmathew1MiniProjectMLThis dataset is designed for training an audio classification model that identifies the type of trash being thrown into a bucket.
The model classifies sounds into the following categories: Metal, Glass, Plastic, Cardboard, and Noise (non-trash-related sounds).
The dataset was recorded and organized as part of an Edge Impulse project to create a system that sorts trash based on sound.
Link to Edge Impulse: https://studio.edgeimpulse.com/public/556872/live
Hindi50math-speech-datasetMathBridge2Hindi50_2MathSpeech_whisper_transcribedmr-collin-hegarty-mathsSub to Tubular Pickaxe
SonicFleetwayMatheusMatheusJennMathSpeech_whisper_transcribed_normalizedAudiosPortuguescantoresFleetwaySonicMATHEUS_FNAFy
