CoolFace
20 results

Podcast

Juancarlosy /eluniversoraro-podcastaudion<1K0 likes4.3k downloads2mo agoHugging FaceWhispering-GPT /lex-fridman-podcast-transcript-audio Dataset Card for "lexFridmanPodcast-transcript-audio" Dataset Summary This dataset is created by applying whisper to the videos of the Youtube channel Lex Fridman Podcast. The dataset was created a medium size whisper model. Languages Language: English Dataset Structure The dataset contains all the transcripts plus the audio of the different videos of Lex Fridman Podcast. Data Fields The dataset is composed by: id: Id of the youtube… See the full description on the dataset page: https://huggingface.co/datasets/Whispering-GPT/lex-fridman-podcast-transcript-audio.audioautomatic-speech-recognitionn<1K0 likes2.1k downloads4y agoHugging FaceReadyAi /5000-podcast-conversations-with-metadata-and-embedding-dataset 🗂️ ReadyAI - 5,000 Podcast Conversations with Metadata and Embedding Dataset ReadyAI, operating subnet 33 on the Bittensor Network is an open-source initiative focused on low-cost, resource-minimal pipelines for structuring raw data for AI applications. This dataset is part of the ReadyAI Conversational Genome Project, leveraging the Bittensor decentralized network. AI runs on structured data — and this dataset bridges the gap between raw conversation transcripts and structured… See the full description on the dataset page: https://huggingface.co/datasets/ReadyAi/5000-podcast-conversations-with-metadata-and-embedding-dataset.text10K<n<100K8 likes1.2k downloads1y agoHugging FaceTTS-AGI /podcast-tokenized-bg3.5-enj5-with-speaker-embeddings podcast-tokenized-bg3.5-enj5-with-speaker-embeddings This dataset extends TTS-AGI/podcast-tokenized-bg3.5-enj5 with speaker embeddings, cosine similarity scores, speaker cluster assignments, and reference-match flags for each sample. What was added Each sample's JSON metadata is augmented with the following fields: Field Type Description target_speaker_embedding list[float] (128-dim) L2-normalized speaker embedding of the target audio… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/podcast-tokenized-bg3.5-enj5-with-speaker-embeddings.text-to-speech0 likes1.1k downloads3mo agoHugging FaceEQ4You /podcastvideos4 likes915 downloads2y agoHugging Facealea-institute /dotgov-podcast-sampleaudion<1K0 likes418 downloads2y agoHugging Face