datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
audio-maister
Intro
This is the dataset used to train AudiomAIster. It is mixed with data from VoiceFixer, preprocessed and re-encoded to FLAC.
In addition, we added sound effects to train the model to extract desirable noise (like picking up an object or a musical beat or melody).
Sound effect credits
https://freesound.org/people/airmedia/sounds/349855/
https://freesound.org/people/UnderlinedDesigns/sounds/191766/
https://freesound.org/people/frankum/sounds/324881/… See the full description on the dataset page: https://huggingface.co/datasets/peterwilli/audio-maister.lds-youth-music-chunked
Dataset Card for "lds-youth-music-chunked"
More Information needed
petit-prince-tts-walloon-male
Li Ptit Prince — Walloon TTS (Male Speaker)
A single-speaker Walloon text-to-speech dataset: short, sentence-level audio/text pairs extracted
from a full-length audiobook reading of Li Ptit Prince (the Walloon translation of Antoine de
Saint-Exupéry's Le Petit Prince / The Little Prince).
Source data
Text: wa.wikisource.org — Li Ptit Prince (Hendschel-Mahin, 2023),
the Walloon translation by Lorint Hendschel, with editorial work and phonetic standardization
by… See the full description on the dataset page: https://huggingface.co/datasets/fdemelo/petit-prince-tts-walloon-male.pet-decoder-audio-fixtures
Pet Decoder Audio Test Fixtures (v1.0) 🧪
This repository contains audio test fixtures used for validating the ingestion pipeline of the Pet Decoder AI application.
These are verified, public domain samples used to test our audio visualization and classification algorithms against known baselines (e.g., Low Frequency vs. High Frequency vocalizations).
Dataset Contents
The dataset consists of 5 reference audio files representing distinct spectral patterns:… See the full description on the dataset page: https://huggingface.co/datasets/petdecoder/pet-decoder-audio-fixtures.lds-youth-music
Dataset Card for "lds-youth-music"
More Information needed
petuhov_lecture_2_20petra-sadouski-poshuki-svaigo-vyraiu-petra-sadouski
Пошукі свайго выраю
Metadata
Author: Пётра Садоўскі
Title: Пошукі свайго выраю
Narrator: Пётра Садоўскі
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/petra-sadouski-poshuki-svaigo-vyraiu-petra-sadouski.swahili-tts-roselanguage:
- sw
license: apache-2.0
task_categories:
- text-to-speech
task_ids:
- speech-synthesis
tags:
- swahili
- tts
- speech-synthesis
- audio
- kiswahili
- rose
- female-voice
---
# Swahili TTS Dataset - Rose
A Swahili text-to-speech dataset featuring 500 high-quality audio recordings by female speaker Rose.
## Dataset Description
A comprehensive Swahili text-to-speech dataset featuring **500 high-quality audio recordings** by a female speaker named Rose.
### Dataset Summary
-… See the full description on the dataset page: https://huggingface.co/datasets/PeterPatrick/swahili-tts-rose.audio-maister-val
Dataset Card for "audio-maister-val"
More Information needed
audiotestsven-nurdkvist-petsan-i-findus-ganna-khitryk-i-paval-kharlanchuk
Пэтсан і Фіндус
Metadata
Author: Свэн Нурдквіст
Title: Пэтсан і Фіндус
Narrator: Ганна Хітрык і Павал Харланчук
Source Group: Дзіцячыя
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/sven-nurdkvist-petsan-i-findus-ganna-khitryk-i-paval-kharlanchuk.peteswahili-voice-commandspeter01petryPeter_Glantz_Ivashchenkopetra-sadouski-moi-shybolet-autabiiagrafichnyia-arabeski-petra-sadouski
Мой шыболет. Аўтабіяграфічныя арабэскі
Metadata
Author: Пётра Садоўскі
Title: Мой шыболет. Аўтабіяграфічныя арабэскі
Narrator: Пётра Садоўскі
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/petra-sadouski-moi-shybolet-autabiiagrafichnyia-arabeski-petra-sadouski.peterpeterpetergriffinclassicPetersonAdrianoTobiaspetersteeledatasetvoice_Peter_Capaldivoice_Steve_PetersDailyTalkContiguous-Peter-Griffin
Peter Griffin's Emotion-Tagged DailyTalk Dataset
Hehehehehe! Hey Lois, look! I made a dataset! This is an emotion-tagged version of that DailyTalk thing, but better because it's got all sorts of feelings and stuff. It's freakin' sweet for making computers talk like they've had too many Pawtucket Patriots or just found out Meg is home.
What is this thing?
This dataset has a bunch of people talking, but we tagged 'em with emotions. It's like when I'm happy because it's… See the full description on the dataset page: https://huggingface.co/datasets/MysticKit/DailyTalkContiguous-Peter-Griffin.jorgevozpetite-hostel-tts
