shunyalabs/asturian-speech-dataset
Dataset Card Dataset Description This dataset contains parts of data from Google FLEURS (Few-shot Learning Evaluation of Universal Representations of Speech). Dataset Summary The dataset includes audio recordings sampled at 16kHz along with their corresponding transcripts. It is split into training, validation, and test sets for speech recognition and related tasks. Dataset Structure Data Fields audio: An audio file… See the full description on the dataset page: https://huggingface.co/datasets/shunyalabs/asturian-speech-dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face