TBOGamer22/BrahuiSpeech-70H-V2
BrahuiSpeech-70H V2 BrahuiSpeech-70H V2 is an approximately 70-hour automatic speech recognition dataset containing 15,626 audio-transcription pairs and 69 hours, 32 minutes, 12 seconds of real-world Brahui (Brahvi) speech. Brahui (brh) is a low-resource Dravidian language spoken primarily in Balochistan, Pakistan. The dataset covers naturally occurring speech across varied speakers, speaking styles, media domains, and acoustic conditions. Transcriptions use the Perso-Arabic… See the full description on the dataset page: https://huggingface.co/datasets/TBOGamer22/BrahuiSpeech-70H-V2.
Update README.md
Add BrahuiSpeech-70H V2 dataset card
Publish BrahuiSpeech-70H V2 audio and adjudicated transcripts
Update README.md
Add V2 release statistics
Add BrahuiSpeech-70H V2 dataset card
Publish BrahuiSpeech-70H V2 audio and adjudicated transcripts
initial commit
