dineshkarki/nepali-textbooks-math-grade10
Nepali Textbook Pretraining Corpus — Sample (Grade 10 Math) This dataset contains OCR-extracted, chapter-first, chunked text from Nepali school textbooks. Schema id — unique segment id source — book/source string book_id — filename-derived id (if present) subject — subject label (e.g., "math") grade — class/grade chapter_index — chapter number chapter_title — chapter/unit name segment_index — index within chapter text — content chunk tokens_approx — rough token… See the full description on the dataset page: https://huggingface.co/datasets/dineshkarki/nepali-textbooks-math-grade10.
This repository belongs to dineshkarki on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
