datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ja_asr.jsut_basic5000VoidLinuxISOSJSUT-basic5000
JSUT (Japanese Speech Corpus) - Test Subset
A test subset of the JSUT corpus containing 500 Japanese utterances from the basic5000 dataset (BASIC5000_4501-5000).
Dataset Structure
jsut_ver1.1/
└── basic5000/
├── wav/ # WAV audio files (500 files, 48kHz)
├── transcript_utf8.txt # Transcriptions
└── recording_info.txt # Recording dates
File Formats
transcript_utf8.txt
BASIC5000_4501:だが、エーアイセンター稼動を快く思わない...… See the full description on the dataset page: https://huggingface.co/datasets/FluidInference/JSUT-basic5000.pt_basicsphonetically diverse standalone words, letters, diphtongs and basic greetings
JapanesePitchAccentRecognition_JSUT-basic5000SentenceJapanesePitchAccentRecognition_JSUT-basic5000PhraseBasic-Chord-Ukulele-Data
Basic Chord Ukulele Data
© 2024 by Fasai Rakphakdee
This dataset is licensed under the Creative Commons Attribution-NonCommercial 4.0 International License.To view a copy of this license, visit Creative Commons.
Description
The Basic Chord Ukulele Data contains audio samples and metadata for basic ukulele chords. It is designed to support musicians, educators, and learners who want high-quality ukulele chord sounds and detailed chord information.
Modality
01 :… See the full description on the dataset page: https://huggingface.co/datasets/FasaiRakphakdee/Basic-Chord-Ukulele-Data.
