dubbing
Datasets
All datasets matching “dubbing”vaja-thai
Vaja-Thai (วาจา) — Combined Thai TTS Dataset
A unified, quality-filtered Thai speech dataset combining multiple sources for
Text-to-Speech (TTS) research. All audio is resampled to 24 kHz WAV format.
Dataset Summary
Metric
Value
Total samples
289,916
Total hours
554.6h
Sampling rate
24,000 Hz
Format
WAV 16-bit PCM
Language
Thai (ภาษาไทย)
Sources
Source
Samples
Hours
License
Description
tsync2
1,823
3.7h
CC-BY-NC-SA-3.0
NECTEC… See the full description on the dataset page: https://huggingface.co/datasets/dubbing-ai/vaja-thai.ai-dubbing-videoarticle-14-universal-live-dubbing
Universal Live Dubbing doesn't phone home. It doesn't need to.
Universal Live Dubbing: Real-Time Whisper Transcription and Neural Translation for Web Video
The Problem
The multilingual web presents a fundamental accessibility barrier: video content produced in one language remains inaccessible to speakers of other languages until subtitles or dubbed audio are produced?a process that is expensive, slow, and incomplete. This paper presents Universal Live Dubbing, a… See the full description on the dataset page: https://huggingface.co/datasets/kleinnner/article-14-universal-live-dubbing.article-14-universal-live-dubbing
Universal Live Dubbing doesn't phone home. It doesn't need to.
Universal Live Dubbing: Real-Time Whisper Transcription and Neural Translation for Web Video
The Problem
The multilingual web presents a fundamental accessibility barrier: video content produced in one language remains inaccessible to speakers of other languages until subtitles or dubbed audio are produced?a process that is expensive, slow, and incomplete. This paper presents Universal Live Dubbing, a… See the full description on the dataset page: https://huggingface.co/datasets/Anticloud/article-14-universal-live-dubbing.
