datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
FeatureCoding-KimiAudio@misc{pang2026largemodelfeaturecoding,
title={Towards Large Model Feature Coding},
author={Youwei Pang and Changsheng Gao and Dong Liu and Huchuan Lu and Weisi Lin},
year={2026},
eprint={2605.24025},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2605.24025},
}
FeatureCoding-KimiAudio-AfterCodec@misc{pang2026largemodelfeaturecoding,
title={Towards Large Model Feature Coding},
author={Youwei Pang and Changsheng Gao and Dong Liu and Huchuan Lu and Weisi Lin},
year={2026},
eprint={2605.24025},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2605.24025},
}
Kimi-Audio-GenTest
Kimi-Audio-Generation-Testset
Dataset Description
Summary: This dataset is designed to benchmark and evaluate the conversational capabilities of audio-based dialogue models. It consists of a collection of audio files containing various instructions and conversational prompts. The primary goal is to assess a model's ability to generate not just relevant, but also appropriately styled audio responses.
Specifically, the dataset targets the model's proficiency in:… See the full description on the dataset page: https://huggingface.co/datasets/moonshotai/Kimi-Audio-GenTest.kimi-audio-dpokimi_audio_datasetKimi-Audio-GenTestbbx
