CoolFace
Datasetpublic

firdavsus/Gemma-4-E4B_hidden_to_audio_tokens-2.0

Gemma-4 S2S Alignment Dataset (Multilingual) – Version 2.0 (260K) This dataset is specifically engineered to train a lightweight, low-latency Hidden-to-Speech (H2S) alignment model. By capturing the raw, abstract semantic representations from the 33rd hidden states of a Text LLM (Gemma-4 8B) and mapping them directly onto quantized discrete audio streams, this dataset bypasses traditional text generation bottlenecks to establish native Speech-to-Speech (S2S) processing… See the full description on the dataset page: https://huggingface.co/datasets/firdavsus/Gemma-4-E4B_hidden_to_audio_tokens-2.0.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
0likes42downloads
settings

This repository belongs to firdavsus on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameGemma-4-E4B_hidden_to_audio_tokens-2.0
visibilitypublic
licenceapache-2.0
gatedno
ownerfirdavsus
Account settings
firdavsus/Gemma-4-E4B_hidden_to_audio_tokens-2.0 · CoolFace