datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vital
Vital PresetShare Renders
Rendered Vital presets scraped from PresetShare.
sample rate: 22050
render duration: 6.0s
MIDI note: 72 (C4)
note duration: 5.0s
velocity: 100
Files are organized under by_type/<sound-type>/<preset-id>_<name>/ with:
preset.vital
preview.mp3
vital-render.wav
metadata.json
See manifest.jsonl and summary.json for run metadata.
meow-10k
Dataset Card for Meow-10K
Meow-10K is a high-fidelity, synchronized quad-modal dataset comprising 10,000 feline samples. It is the primary training corpus for Meow-Omni 1, designed to facilitate deep intention reasoning in computational ethology.
Dataset Summary
Meow-10K provides the first large-scale training foundation for Multimodal Large Language Models (MLLMs) to learn the causal relationships between external behaviours and internal physiological states. By… See the full description on the dataset page: https://huggingface.co/datasets/Duckyle/meow-10k.Bai-cho-NgocDUCK04082026duck-25-04-part2
