datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
samromur_childrenThe Samrómur Children corpus contains more than 137000 validated speech-recordings uttered by Icelandic children.children-story-datasetAudio-Children-Stories-CollectionAudio Chidren Stories Collection
This dataset has 600 audio files in .mp3 format. This has been created using my existing dataset Children-Stories-Collection.
I have used only first 600 stories for creating this audio dataset.
You can use this for training and research purpose.
Thank you for your love & support.
Audio-Children-Stories-Collection-LargeAudio Chidren Stories Collection Large
This dataset has 5600++ audio files in .mp3 format. This has been created using my existing dataset Children-Stories-Collection.
I have used first 5600++ stories from Children-Stories-1-Final.json file for creating this audio dataset.
You can use this for training and research purpose.
Thank you for your love & support.
Children_Counsel
아동·청소년 상담 데이터셋 (Children Counseling Dataset)
This dataset contains counseling data for children and adolescents, including both audio recordings and transcriptions.
Dataset Structure
The dataset is organized as follows:
audio/: Contains the audio recordings of counseling sessions in MP3 format
data/: Contains JSON files with transcriptions and metadata for each session
Usage
This dataset can be used for:
Training speech recognition models for counseling… See the full description on the dataset page: https://huggingface.co/datasets/ironDong/Children_Counsel.ChildrenSongTranscriptVerification_CSDChildrenSLIchildren-phoneme-filteredchildren-phoneme-cleaned-1.5pct-removed
