CoolFace
20 results

p1

VideoUFO /ResearchData_P10 likes19k downloads25d agoHugging Faceorionweller /mmBERT-pretrain-p1-fineweb2-langs mmBERT Pre-training Data P1 Phase 1 of 3: Diverse multilingual pre-training data mixture (trained for 2.3T tokens) used to train the mmBERT model suite. NOTE: this is only P1 of the pre-training data due to HF limits, you need to download and combine all three into one folderThis dataset contains the pre-training phase data used to train all mmBERT encoder models. The data is provided in MDS format ready for use with Composer and the ModernBERT training repository.… See the full description on the dataset page: https://huggingface.co/datasets/orionweller/mmBERT-pretrain-p1-fineweb2-langs.fill-mask7 likes5.4k downloads1y agoHugging Facesyvai /p1-segmentsgated DR P1 speech segments Dataset Danish speech clips from DR P1, in mono 16 kHz OGG/Opus, with verbatim text, timing, and speaker metadata. Transcript text and speaker attribution may contain automated errors. Source The recordings cover roughly 2006–2022 and come from DR P1 recordings in kb.dk’s DR archive. Audio is sourced through the pinned syvai/p1 revision 449b9c2294026df6d0d37538f279fdec03f565ff. Transcripts were generated with ElevenLabs… See the full description on the dataset page: https://huggingface.co/datasets/syvai/p1-segments.audioautomatic-speech-recognition1M<n<10M4 likes2.4k downloads5d agoHugging Facep1k0 /HBGimagen<1K0 likes2k downloads12d agoHugging FaceTaoMagnet /sn38-r11-p1textn<1K0 likes1.2k downloads14d agoHugging Facesyvai /p1 DR P1 Audio Archive Danish public radio (DR) P1 audio recordings sourced from the kb.dk DR-arkivet (Royal Danish Library DR archive), covering roughly 2006–2022. Format Audio: Opus, 24 kbps, mono, in OGG container (transcoded from DR's mp3 archive) Parquet shards (~500 items each), small row groups for streaming compatibility Sortable by year / month / start_time Schema Each row is one broadcast item with the full audio bytes inline plus rich… See the full description on the dataset page: https://huggingface.co/datasets/syvai/p1.audioautomatic-speech-recognition100K<n<1M2 likes888 downloads4mo agoHugging Face