Willy030125/ambient_noise_audio
Ambient Noise Collection dataset This dataset is useful for reducing ASR model Hallucinations (especially Whisper), by default Whisper often transcribing non-speech audio as hallucination transcriptions. This dataset attempts to improve ASR model that have hallucinations on non-speech audio. Collected several audio from source: https://www.kaggle.com/datasets/nafin59/hospital-ambient-noise https://www.kaggle.com/datasets/solorzano/ambient-noise… See the full description on the dataset page: https://huggingface.co/datasets/Willy030125/ambient_noise_audio.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face