CoolFace
20 results

seq2seq

speech-seq2seq /amiGigaSpeech is an evolving, multi-domain English speech recognition corpus with 10,000 hours of high quality labeled audio suitable for supervised training, and 40,000 hours of total audio suitable for semi-supervised and unsupervised training. Around 40,000 hours of transcribed audio is first collected from audiobooks, podcasts and YouTube, covering both read and spontaneous speaking styles, and a variety of topics, such as arts, science, sports, etc. A new forced alignment and segmentation pipeline is proposed to create sentence segments suitable for speech recognition training, and to filter out segments with low-quality transcription. For system training, GigaSpeech provides five subsets of different sizes, 10h, 250h, 1000h, 2500h, and 10000h. For our 10,000-hour XL training subset, we cap the word error rate at 4% during the filtering/validation stage, and for all our other smaller training subsets, we cap it at 0%. The DEV and TEST evaluation sets, on the other hand, are re-processed by professional human transcribers to ensure high transcription quality.0 likes2.1k downloads4y agoHugging Faceaklein4 /seq2seq-mixed-pretraining-SmolLM2tabular100M<n<1B1 likes1.6k downloads8mo agoHugging Faceclosji /seq2seq-glue Dataset Card for "seq2seq-glue" More Information needed text1M<n<10M0 likes197 downloads3y agoHugging Facespeech-seq2seq /ami-ihm-kaldi-processed0 likes161 downloads4y agoHugging FaceYuhthe /phoner_seq2seq Dataset Card for "phoner_seq2seq" More Information needed text10K<n<100K0 likes80 downloads3y agoHugging FaceErfanMoosaviMonazzah /esnli-seq2seqtext100K<n<1M0 likes80 downloads3y agoHugging Face