CoolFace
Datasetpublic

Noothi/huberman-lab-transcripts

Huberman Lab Transcript Dataset Cleaned English transcripts from 438 videos published on the Huberman Lab YouTube channel. Dataset 438 videos 9,833 transcript chunks ~114 million characters JSONL format Each record contains: ext ideo_id itle Processing The transcripts were collected from YouTube captions and processed by normalizing whitespace, removing common caption artifacts, removing repeated words, splitting into coherent chunks, and… See the full description on the dataset page: https://huggingface.co/datasets/Noothi/huberman-lab-transcripts.

sourceHugging Faceupdated 18d agoView on Hugging Face
0likes78downloads

Noothi/huberman-lab-transcripts · main · files are served by the source, never re-hosted here