datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nasle-mana-clean-chunked-30s-avasanj
Nasl-e-Mana Clean Persian Speech — corrected 30-second chunks
Corrected, provenance-preserving audio chunks collected from the Nasl-e-Mana magazine website, generated on 2026-08-30. This release supersedes the earlier unreliable proportional-mapping chunk export; that older release was not used here.
Splits
Split
Rows
Audio
Columns
labeled
4,981
41.41 hours
audio, label
to_transcribe
11,127
92.72 hours
audio
The labeled split contains the… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/nasle-mana-clean-chunked-30s-avasanj.dia-AvaAvd-test
AVA-AVD — test split (audio-visual speaker diarization in the wild)
Copie du split test d'AVA-AVD (Xu et al., ACM MM 2022), un corpus de
diarisation dans des conditions "in the wild" construit sur AVA Active Speaker.
Chaque clip (~5 min) est extrait des vidéos AVA à des offsets précis définis
par le repo officiel, puis l'audio est re-synchronisé : les timestamps des
RTTMs ont été recalculés en temps clip-local (soustrait min_start du
RTTM d'origine) pour être directement utilisables… See the full description on the dataset page: https://huggingface.co/datasets/ggfox00000/dia-AvaAvd-test.
