CoolFace
Datasetpublicgated

milanakdj/nepali-audio-reserve-r6

Nepali two-speaker conversation chunks ~6680.8 h of Nepali speech at 48 kHz. Two speakers per clip, ~5 minute diarized chunks. A backup, not a release: the transcripts are machine-generated, and none of this audio passed the quality gate that produced our training corpus. Derived from third-party audio whose rights holders did not grant redistribution. The hour count is language-dominant, not monolingual: a chunk labelled Nepali can carry substantial English or Hindi. lang_sec… See the full description on the dataset page: https://huggingface.co/datasets/milanakdj/nepali-audio-reserve-r6.

sourceHugging Faceotherupdated 16d agoView on Hugging Face
0likes341downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.