CoolFace
Datasetpublicgated

InfoBayAI/Marathi-STEM-Textbook-Dataset

Dataset Description: This dataset is a large-scale collection of Marathi STEM textbook data, containing 173 books and 7.81 million words, designed to support the development and training of advanced NLP systems and AI models for scientific understanding, problem-solving, and concept learning in Marathi. Full Dataset Overview This dataset is part of a large-scale multilingual educational corpus containing over 3+ billion words across 5,000+ subjects, supported by interwoven images for deeper… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Marathi-STEM-Textbook-Dataset.

sourceHugging Facecc-by-4.0updated 7d agoView on Hugging Face
0likes19downloads
.gitattributesDownload Raw Back to root

This repository is gated, so its file contents are only served once you have accepted the publisher's terms at Hugging Face. Open it at the source above.