CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01rlabz /mwanamke_moshi Swahili Moshi Fine-tuning Dataset (mwanamke) Overview This dataset prepares Swahili conversational audio for fine-tuning Moshika (the female-voice Moshi variant) using kyutai-labs/moshi-finetune. It builds on the stereo, speaker-separated audio chunks from rlabz/qsuperposition_mwanamke and adds the .jsonl index and per-file .json transcripts that moshi-finetune requires for training. Source Data Origin: rlabz/qsuperposition_mwanamke — stereo… See the full description on the dataset page: https://huggingface.co/datasets/rlabz/mwanamke_moshi.audiotext-to-speech0 likes120 downloads29d agoHugging Face02Nobhpadol /pre-moshi-madam0 likes114 downloads5d agoHugging Face03kalbin /moshi-on-policy-prompts-3kBased on https://huggingface.co/datasets/yhytoto12/behavior-sd audio1K<n<10K0 likes77 downloads5mo agoHugging Face04kalbin /moshi-on-policy-dpo-margin3audio1K<n<10K0 likes68 downloads5mo agoHugging Face05Ethan615 /moshi-zh-domain_wavfile0 likes54 downloads4mo agoHugging Face06moshiko123 /daily-life-hacks-grocery-nutrition-per-dollar Daily Life Hacks: Grocery Nutrition per Dollar This dataset contains two public CSV files ranking grocery foods by nutrition per dollar: fiber-per-dollar-2026.csv — fiber per dollar protein-per-dollar-2026.csv — protein per dollar Methodology and disclosure Nutrition values come from USDA FoodData Central, and prices are US prices from July 2026. This is an independent study and is not endorsed by USDA. Read the companion articles: Cheapest high-fiber foods… See the full description on the dataset page: https://huggingface.co/datasets/moshiko123/daily-life-hacks-grocery-nutrition-per-dollar.0 likes47 downloads18d agoHugging Face07robinwitch /xx_beat_arkit_moshi_2025_07_20_30fps_attntext1K<n<10K0 likes46 downloads1y agoHugging Face08MoshinAli /greenobe-bd-resultsdocumentn<1K0 likes44 downloads27d agoHugging Face09robinwitch /zeroeggs_moshi_2025_06_06_30fps_conv4text1K<n<10K0 likes29 downloads1y agoHugging Face10abrarfahim /moshi-tool-audio Moshi Tool-Calling — Audio-Grounded Dataset Audio-grounded data teaching Moshi / PersonaPlex to emit tool-call special tokens in its inner monologue when it hears a request — and to stay quiet otherwise (listening/idle frames are trained to PAD). Each row is a code tensor codes[17, T] at 12.5 Hz: rows stream content 0 text monologue PAD while listening/idle, `< 1:9 Moshi audio silence 9:17 user audio the spoken question (edge-tts), Mimi-encoded mask=1 marks… See the full description on the dataset page: https://huggingface.co/datasets/abrarfahim/moshi-tool-audio.audioautomatic-speech-recognition1K<n<10K0 likes29 downloads3mo agoHugging Face11kalbin /moshi-paired-prefaudio1K<n<10K0 likes27 downloads4mo agoHugging Face12svjack /moshimoshi_ai_girls_zh_captionedimagen<1K0 likes24 downloads9mo agoHugging Face13chtugha /small-german-medical-dialogue-dataset-for-moshi Small german dialogue dataset This dataset contains 500 completely made up medical phonecall dialogues between patients and a GP's office. Dataset Details Dataset Description 500 made up phonecalls that were first created with AI as text. The audio was then created using Openai tts-1-hd and the accurately timestamped transcripts were added. The audio files are formatted like this: Stereo with split channels: Speaker A is on the left channel… See the full description on the dataset page: https://huggingface.co/datasets/chtugha/small-german-medical-dialogue-dataset-for-moshi.audioaudio-text-to-textn<1K0 likes21 downloads4mo agoHugging Face14chiyuanhsiao /moshi_chunk3audion<1K0 likes19 downloads1y agoHugging Face15ak3ra /sunbird-moshi-development-eval-v1audio1K<n<10K0 likes18 downloads1mo agoHugging Face16chiyuanhsiao /eval_moshi_group_1audion<1K0 likes16 downloads10mo agoHugging Face17robinwitch /zeroeggs_moshi_2025_05_29text1K<n<10K0 likes15 downloads1y agoHugging Face18kalbin /moshi-on-policy-dpo-v17-partialaudio1K<n<10K0 likes14 downloads5mo agoHugging Face19chiyuanhsiao /moshi-continuationaudion<1K0 likes13 downloads9mo agoHugging Face20kalbin /moshi-on-policy-dpo-v20-kyutai-smokeaudion<1K0 likes13 downloads4mo agoHugging Face21anthony-wss /moshi_tts_dataset_dummytextn<1K0 likes12 downloads2y agoHugging Face22kalbin /moshi-on-policy-dpo-v20-kyutai-alignedaudio1K<n<10K0 likes11 downloads4mo agoHugging Face23chiyuanhsiao /moshi_chunk23audion<1K0 likes8 downloads1y agoHugging Face24kalbin /moshi-on-policy-dpo-9kaudio1K<n<10K0 likes8 downloads5mo agoHugging Face25kalbin /moshi-on-policy-dpo-tts-v18-fullprompt-smokeaudion<1K0 likes8 downloads4mo agoHugging Face26robinwitch /zeroeggs_moshi_2025_05_26_validtext1K<n<10K0 likes7 downloads1y agoHugging Face27isLucid /moshi-lt-data Moshi-LT Training Data Synthetic Lithuanian audio data for training Moshi voice agents. Contents Split Files Duration Format Monologues 14899 unknownh Mono WAV 24kHz + JSON timestamps Dialogues 1703 unknownh Stereo WAV 24kHz + JSON timestamps Total 16602 unknownh Generation TTS engine: Google Chirp 3: HD (28 Lithuanian voices) Monologues: Wikipedia + CulturaX Lithuanian text Dialogues: LLM-generated scripts (Gemini 2.0 Flash) —… See the full description on the dataset page: https://huggingface.co/datasets/isLucid/moshi-lt-data.audio1K<n<10K0 likes7 downloads7mo agoHugging Face28robinwitch /xx_moshi_2025_07_09_30fps_conv4text1K<n<10K0 likes6 downloads1y agoHugging Face29chiyuanhsiao /moshi_chunk10audion<1K0 likes6 downloads1y agoHugging Face30svjack /moshimoshi_ai_videos videon<1K0 likes5 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.