datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
NEXUS-temporal_hierarchical_multi-modal
NEXUS: Neural Evolution for eXtensible Universal Semantics Dataset
(Temporal Multimodal Slices)
This dataset is a multi-modal, hierarchical, temporal representation derived from HuggingFaceFV/finevideo. It is designed for streaming training where the primary unit is a 10 ms "slice" that aggregates upward into moments (100 ms), seconds (1 s), experiences (10 s), and minutes (60 s).
It is meant to represent an extensible stream of "experience" as there are… See the full description on the dataset page: https://huggingface.co/datasets/Ardea/NEXUS-temporal_hierarchical_multi-modal.Nexus_image6
Bhagavad-Gita_Audio
Dataset Summary
Bhagavad-Gita_TTS is a high-quality, verse-aligned audio dataset of the Bhagavad Gita, designed for Text-to-Speech (TTS), Automatic Speech Recognition (ASR), Sanskrit NLP, and spiritual audio research.
Each shloka is paired with:
Original Sanskrit text
IAST transliteration
Clean High Quality 44.1KHz WAV audio recordings
The dataset is structured to mirror the 18 chapters of the Gita, covering all 701 shlokas… See the full description on the dataset page: https://huggingface.co/datasets/Sarkhan222/Nexus_image6.
