indian-data
llama3.1_finetuned_on_indian_legal_datasetspeech_emotion_wav2vec_indianEnglish_greek_pretrained_balanced_dataser_wav2vec_indianEnglish_greek_pretrained_balanced_dataaugmented_indians_dataset_client2adaption_indian_finance_datasetwav2vec2-Indian_English_Accent_datasetIndian_Financial_News_model_trained_on_reduced_dataindian-legal-compliance-dataset
minuszero-indian-autonomous-driving-dataset-v2
INDUS-AD: Indian Dataset of Unstructured Urban Scenes for Autonomous Driving
Overview
INDUS-AD is the largest publicly released Indian autonomous-driving dataset for end-to-end autonomous-driving research. Its name expands to Indian Dataset of Unstructured Urban Scenes for Autonomous Driving.
This gated dataset is the decoded companion to the Minus Zero Indian Urban Autonomous Driving Dataset. It provides directly usable camera MP4s, normalized sensor tables… See the full description on the dataset page: https://huggingface.co/datasets/gagandeepreehal/minuszero-indian-autonomous-driving-dataset-v2.indian-stocks-comprehensive-fundamentals-dataset
Indian Stocks Comprehensive Fundamentals Dataset
From screener.in | 5673 Stocks | 387.58 MB+ Data | Weekly Updates
Highlights :
Total Number of stocks : 5673
Dataset Size : 387.58 MB
Status :
last updated on Tuesday, 22 Sep 2026 09:05:54 +0000
Usage Notes :
They are stored in 5673 individual files.
[stockname].json means all data related to that stock.
For example, titan.json contains all available fundamental data… See the full description on the dataset page: https://huggingface.co/datasets/AYUSHKHAIRE/indian-stocks-comprehensive-fundamentals-dataset.indian-stock-market-minute-data
🇮🇳 Indian Stock Market Data: Minute & Daily (2000 - 2026)
📌 Overview
This is a high-performance financial dataset containing the historical price history of 2,500+ NSE Stocks and Indices.
The dataset has been sharded and optimized for high-speed training. Instead of thousands of tiny files, it is grouped into large ~1.5GB Parquet shards, making it ideal for fast streaming with the Hugging Face datasets library.
📊 Dataset Stats
Total Rows: ~715 Million… See the full description on the dataset page: https://huggingface.co/datasets/xxparthparekhxx/indian-stock-market-minute-data.indian-road-dataset
🚗 Indian Road Driving Dataset
The Indian Road Driving Dataset is the largest open dataset of annotated Indian road footage, created by ThirdEye Labs. It addresses the critical gap in autonomous driving datasets for Indian road conditions.
🌍 Why Indian Roads?
Indian roads present unique challenges absent from existing datasets (BDD100K, nuScenes, Waymo):
Dense mixed traffic with unpredictable behavior
Auto-rickshaws, cattle, and informal lane usage
Extreme… See the full description on the dataset page: https://huggingface.co/datasets/thirdeyelabs/indian-road-dataset.indian-dish-datasetIndian_Sign_Language_Data.gov_Rencoded
Indian_Sign_Language_Data.gov_Rencoded
Dataset Overview
Dataset name: Indian_Sign_Language_Data.gov_RencodedHugging Face repository: silentone0725/Indian_Sign_Language_Data.gov_RencodedModality: Video (H.265 / HEVC)Total size: ~75 GBOriginal size: ~200 GBLanguage: Indian Sign Language (ISL)License: MIT
This dataset is a re-encoded and curated version of the Indian Sign Language Dictionary originally published on the Government of India Open Data Portal (data.gov.in).… See the full description on the dataset page: https://huggingface.co/datasets/silentone0725/Indian_Sign_Language_Data.gov_Rencoded.
