datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
English_Natural_Conversation_ASR_STT
BoxlyX English Natural Conversation Sample Dataset (ASR/STT)
📌 Overview
This repository contains high-fidelity, studio-recorded English natural conversation samples designed for training and benchmarking advanced Automatic Speech Recognition (ASR) and Speech-to-Text (STT) models.
This dataset is a curated public sample provided by BoxlyX AI Solution, showcasing our end-to-end capabilities in premium audio data generation, multi-speaker recording environment… See the full description on the dataset page: https://huggingface.co/datasets/BoxlyX/English_Natural_Conversation_ASR_STT.ts_asr_test
ts_asr_test:目标说话人 ASR 测试集(manifest-only)
ts_asr_test is a 3,928-clip (~8.7 h) Chinese/English target-speaker ASR test set, released
manifest-only: the repo ships no audio, only an audio-free recipe and a self-contained,
deterministic rebuild script. Bring your own copies of the public source corpora and run
rebuild_ts_asr_test.py to regenerate every clip bit-for-bit.
数据集简介
每条样本由一段目标说话人语音、一段同说话人的注册音频(enrollment),以及若干干扰说话人语音与背景噪声按固定配方混合而成。任务:在给定 enrollment… See the full description on the dataset page: https://huggingface.co/datasets/Boxp/ts_asr_test.
