datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vedic-neural-geometry"""
🕉️ Vedic Neural Geometry
वैदिक ज्ञान आणि आधुनिक Neural Networks, Knowledge Graphs, Geometric Embeddings आणि Hybrid RAG यांचा संगम.
📊 Current Statistics (v1.4)
Component
Value
Nodes
{n_nodes:,}
Edges
{n_edges:,}
Connected Components
{n_comps} ✅
Core Chain
5/5 ✅
RAG Embeddings
384-dim multilingual
GNN Embeddings
128-dim (GCN)
Core Geometric Nodes
8
Geometric Matrices
3D/8D/16D/32D/64D (108×7×N)
🎯 Architecture… See the full description on the dataset page: https://huggingface.co/datasets/kalpesh77/vedic-neural-geometry.ournakshatra-vedic-astrology-core
OurNakshatra Vedic Astrology Core Dataset
Dataset Summary
The ournakshatra-vedic-astrology-core dataset is a highly structured, expert-curated collection of 602 Q&A pairs covering foundational and advanced concepts in Vedic Astrology (Jyotish). It was developed by the team at OurNakshatra to address the severe lack of high-quality, hallucination-free Vedic astrology training data available to the open-source AI community.
Modern language models frequently struggle… See the full description on the dataset page: https://huggingface.co/datasets/OurNakshatra/ournakshatra-vedic-astrology-core.vedic-sanskrit
Dataset Card for "vedic-sanskrit"
More Information needed
vedic-dependency-parsing5vedic-dependency-parsing6vedic-dependency-parsing3vedic-dependency-parsing4vedic-accent-restoration-dataset
Citation
@inproceedings{tsukagoshi-2025-accent-restoration,
title = {Automatic Accent Restoration in Vedic Sanskrit with Neural Language Models},
author = {Tsukagoshi, Yuzuki and Ohmukai, Ikki},
booktitle = {Proceedings of the 1st Workshop on Benchmarks, Harmonization, Annotation, and Standardization for Human-Centric AI in Indian Languages (BHASHA 2025)},
editor = {Bhattacharya, Arnab and Goyal, Pawan and Ghosh, Saptarshi and Ghosh, Kripabandhu},
year =… See the full description on the dataset page: https://huggingface.co/datasets/yzk/vedic-accent-restoration-dataset.vedic-sanskrit-sources
Dataset Card for "vedic-sanskrit-sources"
More Information needed
vedic-dependency-parsingvedic-dependency-parsing13vedic-dependency-parsing12vedic-dependency-parsing2Vedic-Sanskrit-SLP1Vedic Sanskrit texts from GRETIL and TITUS.
They are transliterated in SLP1, Sanskrit Library Phonological Text Encoding Scheme 1 (basic).
vedic-math-problems-cleanedsa_vedic-ud-dsvedic-dependency-parsing8vedic-dependency-parsing7vedic-knowledge-base
Vedic Knowledge Base (Shriyantra OS Architecture)
हा डेटासेट वैदिक ज्ञान, चक्र, मंत्र, पवित्र भूमिती (Sacred Geometry) आणि आधुनिक न्यूरल नेटवर्क आर्किटेक्चर (Artificial Neural Networks) यांना एकत्र जोडणारे एक नाविन्यपूर्ण मॉडेल आहे.
रचना (Structure)
मल्टी-कॉन्फिगरेशन सेटिंग्समुळे तुम्ही प्रत्येक फाईल आता स्वतंत्रपणे लोड करू शकता.
vedic-texts
Vedic Texts Dataset
Dataset Summary
This dataset is a comprehensive collection of Vedic texts from various sources, including:
Rigveda (Samhita)
Sankhayana Brahmana
Satapatha Brahmana
Upanishads
Vedangas
The dataset contains 65,140 verses in total, with 58,626 verses in the training set and 6,514 verses in the test set. Each verse is annotated with metadata including its source, genre, and text-specific structural information.
Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/the-ak-2000/vedic-texts.vedic-dependency-parsing9vedic-dependency-parsing10vedic-dependency-parsing-none2vedic-dependency-parsing-none3Vedic-Sanskrit-ISOVedic Sanskrit texts from GRETIL and TITUS.
They are transliterated in ISO 15919.
vedic-math-problemsvedic-rerankingvedic-dependency-parsing-none4mini-platypusvedic-dependency-parsing11
