CoolFace
Datasetpublic

typhoon-ai/gigaspeech2-typhoon

Gigaspeech2 Typhoon Project page | Paper | GitHub Gigaspeech2 Typhoon is a metadata-only reference dataset for Thai speech recognition benchmarking, specifically designed as an Accuracy Track for evaluating ASR models. The dataset contains 1,000 test samples with audio IDs and human transcriptions derived from the Gigaspeech2 corpus. Each audio_id directly links to the original Gigaspeech2 dataset, allowing users to download the corresponding audio. Dataset Overview… See the full description on the dataset page: https://huggingface.co/datasets/typhoon-ai/gigaspeech2-typhoon.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
1likes64downloads
settings

This repository belongs to typhoon-ai on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namegigaspeech2-typhoon
visibilitypublic
licencecc-by-4.0
gatedno
ownertyphoon-ai
Account settings
typhoon-ai/gigaspeech2-typhoon · CoolFace