reazon-research/reazonspeech
Dataset Card for ReazonSpeech Dataset Summary This dataset contains a diverse set of natural Japanese speech, collected from terrestrial television streams. It contains more than 35000 hours of audio. Paper: ReazonSpeech: A Free and Massive Corpus for Japanese ASR Disclaimer TO USE THIS DATASET, YOU MUST AGREE THAT YOU WILL USE THE DATASET SOLELY FOR THE PURPOSE OF JAPANESE COPYRIGHT ACT ARTICLE 30-4. Dataset Format Audio files are… See the full description on the dataset page: https://huggingface.co/datasets/reazon-research/reazonspeech.
This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.
