Kaousheik/tempo
TEMPO — Temporally-grounded Multi-task Post-training for LALMs Training and evaluation data for TEMPO, a unified large audio-language model that assigns timestamps to events, speakers and sounds across speech, sound and music. Every example is a (audio, question, answer) triple whose answer is text interleaved with atomic timestamp tokens at 0.1 s resolution (<|0.0|>, <|0.1|>, … <|60.0|>), prefixed by a task tag. Structure There is one config per task and the… See the full description on the dataset page: https://huggingface.co/datasets/Kaousheik/tempo.
This repository belongs to Kaousheik on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
