CoolFace
Datasetpublic

ud-nlp/human-robot-conversation-german

Human-Robot Conversation Dataset (German) - 660+ Hours Dataset (German) contains 660+ hours of audio featuring dialogues between AI and a human in German across 20,000 recordings. The dataset supports conversational AI, speech recognition, and human-robot interaction research, with short M4A audio files (up to 2 minutes) and structured metadata for model training. - Get the data Dataset characteristics: Characteristic Data Description Audio of… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/human-robot-conversation-german.

sourceHugging Facecc-by-nc-nd-4.0updated 6mo agoView on Hugging Face
1likes11downloads
Dataset Card

Human-Robot Conversation Dataset (German) - 660+ Hours

Dataset (German) contains 660+ hours of audio featuring dialogues between AI and a human in German across 20,000 recordings. The dataset supports conversational AI, speech recognition, and human-robot interaction research, with short M4A audio files (up to 2 minutes) and structured metadata for model training. - [Get the data](https://unidata.pro/datasets/human-robot-conversation-german/?utm_source=huggingface-nlp&utm_medium=referral&utm_campaign=human-robot-conversation-german)

Dataset characteristics:

CharacteristicData
DescriptionAudio of dialogues between AI and humans in German
Data typesAudio
TasksSpeech Recognition, LLM
Hours of audio660+
Number of sets20,000
LanguageGerman
LabelingMetadata (id, language, format)

📊 Sample dataset available! For full access, contact us to discuss purchase terms.

Dataset structure

  • audio.m4a - audio file
  • metainfo.csv - metadata for the data

🧩 Like the dataset but need different data? We can collect a custom dataset just for you - learn more about our data collection services here

Similar Datasets:

  1. 1.American Speech Recognition Dataset
  2. 2.Speech Emotion Recognition Dataset
  3. 3.Human-Robot Conversation Dataset (English)

🌐 UniData - your trusted data partner. Unique, accurate, thoroughly collected and annotated data designed to fuel your AI/ML success.