datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MM-PreTrain
JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation
[HomePage]
[Paper]
[GitHub]
TL;DR
We introduce JavisGPT, a multimodal LLM that can understand audiovisual inputs and simultaneously generate synchronized sounding videos in a unified model.
We also curate the JavisInst-Omni dataset to facilitate instruction-tuning for comprehension and generation on sounding videos.
📰 News
[2025.12.30] 🚀 We release the training… See the full description on the dataset page: https://huggingface.co/datasets/JavisVerse/MM-PreTrain.JavisInst-Omni
JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation
[HomePage]
[Paper]
[GitHub]
TL;DR
We introduce JavisGPT, a multimodal LLM that can understand audiovisual inputs and simultaneously generate synchronized sounding videos in a unified model.
We also curate the JavisInst-Omni dataset to facilitate instruction-tuning for comprehension and generation on sounding videos.
📰 News
[2025.12.30] 🚀 We release the training… See the full description on the dataset page: https://huggingface.co/datasets/JavisVerse/JavisInst-Omni.SLR35_javaneseSpanishPodcastshigh_quality_spanish_speechjavanese-speech-datasetjavaneseJavisData-Audiodhivehi-shaafiu-speechDhivehi Shaafiu Speech is a single speaker Dhivehi speech dataset created by [Javaabu Pvt. Ltd.](https://javaabu.com).
The dataset contains around 16.5 hrs of text read by professional Maldivian narrator Muhammadh Shaafiu.
The text used for the recordings were text scrapped from various Maldivian news websites.dhivehi-javaabu-speech-parquetJavanese-Speech-Dataset
🎧 Javanese Speech Dataset
The Javanese Speech Dataset is a structured and scalable speech audio dataset designed to provide high-quality audio data for training modern AI and machine learning models. It includes 85 hours of audio data across 585 files, delivered in MP3 and WAV formats, with a total size of 104 MB. This well-balanced audio dataset offers diverse and representative voice data, with 51% female and 49% male speakers, and an age range spanning from 18 to 50+ years. The… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Javanese-Speech-Dataset.fork-google-openslr-javanesedhivehi-majlis-speechDhivehi Majlis Speech is a Dhivehi speech dataset created from data annotated by [Javaabu Pvt. Ltd.](https://javaabu.com).
The dataset contains around 10.5 hrs of speech collected from parliament sessions at The Peoples Majlis of Maldives (Maldivian Parliament) consisting of audio from different MPs from 6 different sessions.azerbaijani-speech-datasetopenslr_java_sundadhivehi-khadheeja-speechDhivehi Khadheeja Speech is a single speaker Dhivehi speech dataset created by [Javaabu Pvt. Ltd.](https://javaabu.com).
The dataset contains around 20 hrs of text read by professional Maldivian narrator Khadheeja Faaz.
The text used for the recordings were text scrapped from various Maldivian news websites.JavierMileikusjankou-mikola-javar-z-kalinaju-kacjaryna-jagorava
Kusjankou Mikola, Javar z kalinaju, Kacjaryna Jagorava
Metadata
Original Title (Cyrillic): Кусянкоў Мікола, Явар з калінаю, Кацярына Ягорава
Transliterated Title: Kusjankou Mikola, Javar z kalinaju, Kacjaryna Jagorava
Audio Files: 93 MP3 files
Format: Belarusian audiobook
Description
This is a Belarusian audiobook dataset containing 93 audio tracks.
License
Please check the original source for licensing information.
corte.zipA0001_S003_0_G0001_G0002_continuations_es
Dataset A0001_S003_0_G0001_G0002 para Ultravox (ES)
Clips: 177
Formato: JSONL + WAV en audio/
Subido automáticamente con upload_to_hf.py.
orpheus-es-datasetcv-mn-workshopjavaneseTrain-sample-1000general_knowledge_geminiscaling_dhivehi_stt
