CoolFace
Datasetpublicgated

milanakdj/nepali-audio-reserve-r6

Nepali two-speaker conversation chunks ~6680.8 h of Nepali speech at 48 kHz. Two speakers per clip, ~5 minute diarized chunks. A backup, not a release: the transcripts are machine-generated, and none of this audio passed the quality gate that produced our training corpus. Derived from third-party audio whose rights holders did not grant redistribution. The hour count is language-dominant, not monolingual: a chunk labelled Nepali can carry substantial English or Hindi. lang_sec… See the full description on the dataset page: https://huggingface.co/datasets/milanakdj/nepali-audio-reserve-r6.

sourceHugging Faceotherupdated 16d agoView on Hugging Face
0likes341downloads

No commit history came back for main. The revision may not exist, or the source declined the request.