prosody
childvox-speechocean762-prosody-whisper-basechildvox-speechocean762-prosody-babyhubertchildvox-speechocean762-prosody-whisper-largeqwen3-4b-prosody-rag-lora-mlxkan-bayashi_jsut_vits_prosodyparakeet-tdt-0.6b-HD-prosodykan-bayashi_tsukuyomi_full_band_vits_prosodykan-bayashi_jvs_tts_finetune_jvs010_jsut_vits_raw_phn_jaconv_pyopenjtalk_prosody_latest
Datasets
All datasets matching “prosody”ProsodyEmoji
The Prosody of Emojis (ACL 2026)
Dataset Summary
Prosodic features such as pitch, timing, and intonation are central to spoken communication, conveying emotion, intent, and discourse structure. In text-based settings, emojis act as visual surrogates that add affective and pragmatic nuance.
This dataset examines how emojis influence prosodic realisation in speech and how listeners interpret prosodic cues to recover emoji meanings. It contains human speech data… See the full description on the dataset page: https://huggingface.co/datasets/GiulioZh/ProsodyEmoji.motherese-prosody-data
Prosody Features for train-clean-100
This repository contains a pickled ProsodyFeatureExtractor object trained on the LibriSpeech train-clean-100 and dev-clean subset.
Contents
Word-level prosodic features
F0, energy, duration, pause, prominence
Extracted using CELEX-based stress localization
Format
.pkl file — can be loaded using pickle.load(open(..., "rb"))
Compatible with JSON serialization
macro_prosody_sample_set
Alexandria Voice Corpus — Multilingual Macro-Prosody Telemetry
Version 1.1 — Replacement release
This pack supersedes the earlier Korean & Hindi two-language release. That release was built on a pipeline with several unresolved quality-gate bugs (documented below). This version corrects all known issues and expands to seven typologically diverse languages.
No audio is included. This is a structured acoustic feature dataset for linguistic research, speech technology, and… See the full description on the dataset page: https://huggingface.co/datasets/moonscape-software/macro_prosody_sample_set.Prosody_Breton
[!NOTE]
Dataset origin: https://cocoon.huma-num.fr/exist/crdo/meta/cocoon-3626dbec-0905-4a59-a6db-ec09054a59f7
Description originale
We collected data from Breton dialects to study their prosody and how they inform the syntax-phonology interface. The protocols comprise elicitations and free corpora from Kerne and Treger. The free corpus files include self-portraits of native Breton speakers in their eighties in 2022. The elicited files consist of the results of a… See the full description on the dataset page: https://huggingface.co/datasets/Bretagne/Prosody_Breton.multilingual_audio_alignments_prosodyGroup_H_Chinese-English-Code-Mixing-Prosody
A Multimodal Dataset of Phonological Shift and Prosodic Reset in Chinese-English Code-Mixing
Abstract
Existing code-mixing corpora primarily rely on text transcripts, lacking the precise acoustic alignments necessary to study prosody at the switch boundary. This dataset provides 3 hours of carefully curated, naturalistic Chinese-English code-mixed speech sourced from diverse social media video content. We utilize a dual-level annotation scheme: manual token-level labeling… See the full description on the dataset page: https://huggingface.co/datasets/hafsamenaz1/Group_H_Chinese-English-Code-Mixing-Prosody.
