datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bangla-10k
Bangla-10K: A Challenging, Metadata-Rich Corpus of Read and Conversational Bengali Speech from India and Bangladesh
Bangla-10K is a 10,070.8-hour Bengali speech corpus with
567,323 recordings from India and Bangladesh. It combines scripted
single-speaker read speech with natural multi-speaker conversations for
Bengali automatic speech recognition (ASR). The paper rounds the corpus scale
to 10,000 hours.
The corpus and its ASR evaluation are described in the anonymous manuscript… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/bangla-10k.mandarin-speech-samples
Mandarin Speech Samples
This sample shows Mandarin Chinese speech with clip-level metadata and preview transcripts. It is meant to help buyers review language fit, recording quality, and sample structure before scoping a larger delivery.
What This Shows
Mandarin speech audio with consistent metadata
Clip-level transcript fields for content review
Language and format signals for procurement review
Dataset Specifications
Field
Value… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/mandarin-speech-samples.korean-speech-samples
Korean Speech Samples
This sample shows Korean contributor speech in a consistent audio format. It is meant to help buyers review recording quality, language coverage, and metadata structure before scoping a larger delivery.
What This Shows
Korean speech recordings from contributor collection workflows
Clip-level metadata for format and review context
Ground-truth transcripts for understanding sample content
Dataset Specifications
Field… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/korean-speech-samples.bengali-multi-speaker-speech-samples
Bengali Speech: Multi-Speaker Samples
This sample shows Bengali multi-speaker speech with aligned ground-truth transcripts. It is meant to help buyers review conversational structure, speaker overlap, transcript quality, and audio consistency before scoping a larger delivery.
What This Shows
Multi-speaker Bengali speech with transcript alignment
Conversation-style audio rather than isolated prompt reading
Metadata that distinguishes language, format, and speaker… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/bengali-multi-speaker-speech-samples.tamil-speech-samples
Tamil Speech Samples
This sample shows Tamil speech with ground-truth transcripts and consistent audio metadata. It is meant to help buyers review language fit, transcript quality, and capture format before scoping a larger delivery.
What This Shows
Tamil speech recordings with paired transcripts
Ground-truth labels at the clip level
Format metadata for review and delivery planning
Dataset Specifications
Field
Value
Modality
Audio… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/tamil-speech-samples.cantonese-speech-samples
Cantonese Speech Samples
This sample shows Cantonese speech with native transcript metadata. It is meant to help buyers review dialect fit, recording quality, and sample structure before requesting broader coverage.
What This Shows
Cantonese speech audio with paired transcript metadata
Language-specific metadata for review and delivery planning
A compact preview of the available sample structure
Dataset Specifications
Field
Value… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/cantonese-speech-samples.french-speech-samples
French Speech Samples
This sample shows French contributor speech paired with source transcripts. It is meant to help buyers review recording quality, transcript alignment, and metadata structure before scoping a larger delivery.
What This Shows
French single-speaker recordings
Transcript alignment from the source dataset
Clip-level metadata for format and review context
Dataset Specifications
Field
Value
Modality
Audio
Language… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/french-speech-samples.urdu-speech-samples
Urdu Speech Samples
This sample shows Urdu speech with transcript alignment and simple audio metadata. It is meant to help buyers review language fit and capture quality before requesting a larger sample or production delivery.
What This Shows
Urdu speech audio with paired text
A compact view of transcript and metadata structure
Audio format signals for procurement review
Dataset Specifications
Field
Value
Modality
Audio
Language
Urdu… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/urdu-speech-samples.hindi-speech-samples
Hindi Speech Samples
This sample shows Hindi contributor speech paired with validated text. It is meant to help buyers review spoken content, transcript alignment, and audio format consistency before scoping a larger delivery.
What This Shows
Single-speaker Hindi recordings from contributor collection workflows
Ground-truth transcript alignment at the clip level
Audio metadata suitable for evaluating format and capture consistency
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/hindi-speech-samples.henan-speech-samples
Henan Speech Samples
This sample shows Henan Chinese speech with native transcript metadata. It is meant to help buyers review regional speech fit, recording quality, and sample structure before requesting broader coverage.
What This Shows
Henan Chinese speech audio with paired transcript metadata
Regional language metadata for review and delivery planning
A compact preview of the available sample structure
Dataset Specifications
Field
Value… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/henan-speech-samples.spanish-speech-samples
Spanish Speech Samples
This sample shows Spanish contributor speech in a consistent audio format. It is meant to help buyers review recording quality, language coverage, and metadata structure before scoping a larger delivery.
What This Shows
Spanish speech samples with clip-level review metadata
Clip-level metadata for format and review context
Ground-truth transcripts for understanding sample content
Dataset Specifications
Field
Value… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/spanish-speech-samples.bengali-speech-samples
Bengali Speech Samples
This sample shows Bengali read and conversational speech with paired transcripts. It is meant to help buyers review spoken content, transcript alignment, and audio consistency before scoping a larger delivery.
What This Shows
Bengali speech across read and conversational styles
Clip-level transcript alignment
Audio metadata that supports format and quality review
Dataset Specifications
Field
Value
Modality
Audio… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/bengali-speech-samples.
