CoolFace
Datasetpublicgated

ai4bharat/Spoken-Tutorial

BhasaAnuvaad: A Speech Translation Dataset for 13 Indian Languages Overview BhasaAnuvaad, is the largest Indic-language AST dataset spanning over 44,400 hours of speech and 17M text segments for 13 of 22 scheduled Indian languages and English. This repository consists of parallel data for Speech Translation from Spoken-Tutorial youtube channel, a subset of BhasaAnuvaad. How to use The datasets library allows you to load and pre-process… See the full description on the dataset page: https://huggingface.co/datasets/ai4bharat/Spoken-Tutorial.

sourceHugging Facecc-by-4.0updated 2y agoView on Hugging Face
2likes42downloads
settings

This repository belongs to ai4bharat on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameSpoken-Tutorial
visibilitypublic
licencecc-by-4.0
gatedyes
ownerai4bharat
Account settings
ai4bharat/Spoken-Tutorial · CoolFace