datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
japanese-roleplay-travel-agency
Japanese Travel Agency Roleplay Dialogue Corpus
Overview
The Japanese Travel Agency Roleplay Dialogue Corpus is a collection of ten Japanese role-play dialogues simulating travel-agency consultations between a staff member and a customer.
The conversations were collected for speech and dialogue research and include two-speaker mixed audio, speaker-separated audio, and manually created, time-aligned ELAN (.eaf) annotation files.
The dialogues cover a variety of… See the full description on the dataset page: https://huggingface.co/datasets/IngCrowd/japanese-roleplay-travel-agency.emotional-roleplay-finetuning-dataset
Artificial Voice Roleplay Dataset
67,491 fully-synthetic speech clips (~184 hours) pairing expressive role-play / character
voice-direction captions with generated audio, across German, English, Spanish, and French
(German-dominant). Rich in exaggerated fantasy/creature voices (orc, goblin, troll, ogre,
zombie, dragon, demon, witch, banshee, imp, fairy, gnome, robot, murloc, harpy, skeleton, ghost,
vampire …) and high-arousal emotional delivery (rage, fear, grief, menace).
Every… See the full description on the dataset page: https://huggingface.co/datasets/laion/emotional-roleplay-finetuning-dataset.USA-accented-role-playing-daily-conversations-stereo
Dataset Card for Synthetic daily conversations - USA accented - stereo wav
This dataset consists of synthetic daily conversations recorded by native U.S. English speakers with authentic American accents.
Dataset Details
Dataset Description
This dataset consists of synthetic daily conversations recorded by native U.S. English speakers with authentic American accents. The dialogues are spoken spontaneously, covering a range of everyday topics such… See the full description on the dataset page: https://huggingface.co/datasets/AIxBlock/USA-accented-role-playing-daily-conversations-stereo.English-role-playing-call-center-convers-different-moodsThis dataset features synthetic call center conversations in English, designed to reflect the diversity and complexity of real-world customer service interactions. It includes a broad range of global English accents and emotional tones, making it ideal for training robust conversational AI systems.
🌍 Accents Included: Indian, British (UK), American (USA), Chinese, and more.
🗣️ Speaker Diversity: Features speakers of different genders, age groups, and ethnic backgrounds, all freelancers based… See the full description on the dataset page: https://huggingface.co/datasets/AIxBlock/English-role-playing-call-center-convers-different-moods.jailbreak_audio_roleplay_test_data_04roleplay_train_datasetroleplay_train_data_donaldtrump
