AIxBlock/English-USA-NY-Boston-AAVE-Audio-with-transcription
This dataset captures spontaneous English conversations from native U.S. speakers across distinct regional and cultural accents, including: π½ New York English π Boston English π€ African American Vernacular English (AAVE) The recordings span three real-life scenarios: General Conversations β informal, everyday discussions between peers. Call Center Simulations β customer-agent style interactions mimicking real support environments. Media Dialogue β scripted reads and semi-spontaneousβ¦ See the full description on the dataset page: https://huggingface.co/datasets/AIxBlock/English-USA-NY-Boston-AAVE-Audio-with-transcription.
This dataset captures spontaneous English conversations from native U.S. speakers across distinct regional and cultural accents, including:
π½ New York English
π Boston English
π€ African American Vernacular English (AAVE)
The recordings span three real-life scenarios:
General Conversations β informal, everyday discussions between peers.
Call Center Simulations β customer-agent style interactions mimicking real support environments.
Media Dialogue β scripted reads and semi-spontaneous commentary to reflect media-style speech (e.g. news, interviews, podcasts).
π£οΈ Speakers: Native U.S. English speakers of diverse ages and genders.
π§ Audio Format: Stereo WAV files with separated left/right channels, ensuring high-quality speaker diarization and signal analysis.
π¬ Speech Style: Naturally spoken, unscripted or semi-scripted where appropriate to reflect authentic cadence, tone, and variation.
π Transcriptions: Each audio file is paired with a carefully verified transcript, QAβd by subject matter experts for accuracy and consistency.
π Usage Restriction: This dataset is provided for research and AI model fine-tuning purposes only. Commercialization and resale are strictly prohibited.
π Brought to you by AIxBlock β a decentralized platform for AI dev and workflow automation.
