datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
phonon-youtube-technical
phonon-youtube-technical
This dataset contains Phonon-authored manifests, labels, term/context indexes,
attribution, and local rebuild scripts for technical YouTube speech data.
It intentionally contains no audio files.
Contents
manifests/: segment timing, source URL, video ID, YouTube-reported license,
and audio SHA-256 references.
labels/: Phonon-authored label queues and teacher metadata.
term_context_indexes/: technical term/context indexes used for analysis… See the full description on the dataset page: https://huggingface.co/datasets/Infatoshi/phonon-youtube-technical.ne-en-codeswitching-asr-technical-interview
Dataset Summary
This dataset contains audio recordings and text transcripts of Nepali-English code-switched speech in the context of technical interviews. It is specifically designed to handle the linguistic complexities of Nepali software engineers, developers, and IT professionals who frequently mix English technical terminology (e.g., AWS, S3 lifecycle policies, RAG pipelines, VPC peering) with conversational Nepali grammar.
It is an excellent resource for fine-tuning ASR models… See the full description on the dataset page: https://huggingface.co/datasets/devrahulbanjara/ne-en-codeswitching-asr-technical-interview.
