CoolFace
13 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sanjuhs /audio_to_blendshapes_maintextn<1K0 likes43 downloads1y agoHugging Face02firdavsus /Gemma-4-E4B_hidden_to_audio_tokens-2.0 Gemma-4 S2S Alignment Dataset (Multilingual) – Version 2.0 (260K) This dataset is specifically engineered to train a lightweight, low-latency Hidden-to-Speech (H2S) alignment model. By capturing the raw, abstract semantic representations from the 33rd hidden states of a Text LLM (Gemma-4 8B) and mapping them directly onto quantized discrete audio streams, this dataset bypasses traditional text generation bottlenecks to establish native Speech-to-Speech (S2S) processing… See the full description on the dataset page: https://huggingface.co/datasets/firdavsus/Gemma-4-E4B_hidden_to_audio_tokens-2.0.text100K<n<1M0 likes43 downloads3mo agoHugging Face03Titung /tibetan-to-english-audio-dataset Tibetan to English Audio Dataset Dataset Description A Tibetan speech recognition dataset with transcriptions and English translations containing 1,642 audio samples. Dataset Summary This dataset contains Tibetan speech recordings with: Tibetan transcriptions in native script English translations High-quality audio files in WAV format Total Samples: 1,178Total Size: ~1.1 GBAudio Format: WAV Languages Source Language: Tibetan (བོད་སྐད་) Target… See the full description on the dataset page: https://huggingface.co/datasets/Titung/tibetan-to-english-audio-dataset.audioautomatic-speech-recognition1K<n<10K0 likes33 downloads8mo agoHugging Face04Titung /tibetan-audio-to-english-fixed-filtered Tibetan audio translation Dataset Dataset Description Tibetan audio translation Dataset Dataset Summary This dataset contains 6,366 audio samples with corresponding transcriptions, totaling approximately 15.8 hours of audio. Languages The dataset is in EN (Language code: en). Dataset Structure Data Fields audio: An audio object containing: path: Path to the audio file (if applicable) array: Audio waveform as a numpy array… See the full description on the dataset page: https://huggingface.co/datasets/Titung/tibetan-audio-to-english-fixed-filtered.audioautomatic-speech-recognition1K<n<10K0 likes27 downloads8mo agoHugging Face05sanjuhs /audio_to_blendshapes_testtextn<1K0 likes22 downloads1y agoHugging Face06firdavsus /Gemma-4-E4B_hidden_to_audio_tokens Gemma-4 S2S Alignment Dataset (Multilingual) This dataset is specifically engineered to train a lightweight, low-latency Hidden-to-Speech (H2S) alignment model. By capturing the raw, abstract semantic representations from the last hidden states of a Text LLM (Gemma-4 8B) and mapping them directly onto quantized discrete audio streams, we can bypass traditional text generation bottlenecks to establish native Speech-to-Speech (S2S) processing pipelines. Key Conceptual… See the full description on the dataset page: https://huggingface.co/datasets/firdavsus/Gemma-4-E4B_hidden_to_audio_tokens.text100K<n<1M0 likes12 downloads3mo agoHugging Face07LeroyDyer /Spectrogram_Audio_text_to_Base64imagen<1K2 likes10 downloads2y agoHugging Face08LeroyDyer /Bass_Audio_text_to_Base64imagen<1K1 likes6 downloads2y agoHugging Face09DebottamTalapatra /audio-to-text-datasetaudion<1K0 likes6 downloads8mo agoHugging Face10Anuragleo67 /audio-to-image-sample-knowledge-base Sample Knowledge Base Dataset This folder contains a small starter dataset for the audio-to-image retrieval project. It has 15 educational diagram images and metadata that can be used to build a Pinecone vector index. Files data/sample_knowledge_base/ train/ images/ *.png metadata.csv README.md The images are generated educational diagrams for machine learning topics. What Each Image Record Needs Each image should have these fields:… See the full description on the dataset page: https://huggingface.co/datasets/Anuragleo67/audio-to-image-sample-knowledge-base.imagen<1K0 likes5 downloads5mo agoHugging Face11avdeep /sanskrit_audio_dataset_under_30_taged_meta_to_text_from_edgetabularn<1K0 likes4 downloads2y agoHugging Face12SamagraDataGov /sample_text_to_audio_audiotextn<1K0 likes3 downloads2y agoHugging Face13Sarvesh2003 /sample_text_to_audio_audiotextn<1K0 likes3 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.