datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dahih-tts2-demucs-cleanedpersonaplex-finetuning-pharma-data-sample
PersonaPlex Finetuning — Pharma Data Sample
A 10-example slice of the synthetic patient-support / medication
adherence dataset used to train
demegire/personaplex-finetune-pharma.
The on-disk layout below is exactly what the trainer in
emotion-machine-org/personaplex-finetune
consumes — use this as a template when building your own.
Split: 8 train / 2 eval (mirrors the upstream 2003 / 20 split at
sample scale).
Layout
.
├── adhery_v2.jsonl # master… See the full description on the dataset page: https://huggingface.co/datasets/demegire/personaplex-finetuning-pharma-data-sample.food-asr-tw-demovoice-demo
Multilingual TTS demo — 10 languages of Vietnam and Cambodia
A self-contained Gradio app. Clone the folder, install the requirements, run it.
pip install -r requirements.txt
python -u app.py
Everything resolves relative to app.py, so no paths need editing.
Languages
Code
Language
Code
Language
km
Khmer
tyz
Tay-Nung
blt
Tai Dam
ium
Dao (Iu Mien)
rad
Ede
kpm
Kho
jra
Jarai
cma
Mnong
bdq
Bana
cjm
Cham
blt is Tai Dam, a Tai language of Vietnam… See the full description on the dataset page: https://huggingface.co/datasets/shadwl/voice-demo.
