soda-research/discrete-audio-isoflop-3e18-375M-d896-L9-B8-cb10b5
016
Discrete Audio IsoFLOP Model (discrete-audio-isoflop-3e18-375M-d896-L9-B8-cb10b5)
A suite of discrete audio models trained for our IsoFLOP study as part of SODA, which is a unified next-token prediction on interleaved semantic, acoustic, and text tokens.
๐ฅค Project Page: https://soda-audio.github.io
For full usage instructions (e.g., inference code), and more information, please refer to the [SODA-4B-base](https://huggingface.co/soda-research/soda-4b-base) model card.
The details for this particular model is as follows:
compute_budget: 3e18param_count(non-embedding): 375Mhidden_dim: 896num_layers: 9batch_size: 8training_step: 48960hash_key: cb10b5
๐ WandB: https://wandb.ai/potsawee/marin/groups/IsoFlop/workspace
