tannhoo06/ViMD_preprocessing
ViMD Truncated 10s — 16kHz Preprocessed version of ViMD (Nguyen et al., EMNLP 2024) for Dialect Identification. Preprocessing applied Resample: 44.1kHz (original) → 16kHz, mono Truncate: only the FIRST 10 SECONDS of each audio are kept (files shorter than 10s are kept intact). 1 original file = 1 sample. This follows the truncation strategy of Lu et al. (2020), NOT chunking. Splits: original ViMD train/valid/test kept unchanged (speaker-exclusive).… See the full description on the dataset page: https://huggingface.co/datasets/tannhoo06/ViMD_preprocessing.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face