Podcast
Datasets
All datasets matching “Podcast”eluniversoraro-podcastlex-fridman-podcast-transcript-audio
Dataset Card for "lexFridmanPodcast-transcript-audio"
Dataset Summary
This dataset is created by applying whisper to the videos of the Youtube channel Lex Fridman Podcast. The dataset was created a medium size whisper model.
Languages
Language: English
Dataset Structure
The dataset contains all the transcripts plus the audio of the different videos of Lex Fridman Podcast.
Data Fields
The dataset is composed by:
id: Id of the youtube… See the full description on the dataset page: https://huggingface.co/datasets/Whispering-GPT/lex-fridman-podcast-transcript-audio.5000-podcast-conversations-with-metadata-and-embedding-dataset
🗂️ ReadyAI - 5,000 Podcast Conversations with Metadata and Embedding Dataset
ReadyAI, operating subnet 33 on the Bittensor Network is an open-source initiative focused on low-cost, resource-minimal pipelines for structuring raw data for AI applications.
This dataset is part of the ReadyAI Conversational Genome Project, leveraging the Bittensor decentralized network.
AI runs on structured data — and this dataset bridges the gap between raw conversation transcripts and structured… See the full description on the dataset page: https://huggingface.co/datasets/ReadyAi/5000-podcast-conversations-with-metadata-and-embedding-dataset.podcast-tokenized-bg3.5-enj5-with-speaker-embeddings
podcast-tokenized-bg3.5-enj5-with-speaker-embeddings
This dataset extends TTS-AGI/podcast-tokenized-bg3.5-enj5 with speaker embeddings, cosine similarity scores, speaker cluster assignments, and reference-match flags for each sample.
What was added
Each sample's JSON metadata is augmented with the following fields:
Field
Type
Description
target_speaker_embedding
list[float] (128-dim)
L2-normalized speaker embedding of the target audio… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/podcast-tokenized-bg3.5-enj5-with-speaker-embeddings.podcastvideosdotgov-podcast-sample
