kijjjj/audio_data_russian_annotated
Dataset Audio Russian Annotated This is a dataset with Russian annotated audio data, split into train for tasks like text-to-speech, speech recognition, and speaker identification. Features text: Audio transcription (string). speaker_name: Speaker identifier (string). audio: Audio file. utterance_pitch_mean: The average pitch of the speech utterance (float64). utterance_pitch_std: The standard deviation of pitch, representing variability in intonation (float64)… See the full description on the dataset page: https://huggingface.co/datasets/kijjjj/audio_data_russian_annotated.
2175
