CoolFace
Datasetpublic

dineshkarki/nepali-textbooks-math-grade10

Nepali Textbook Pretraining Corpus — Sample (Grade 10 Math) This dataset contains OCR-extracted, chapter-first, chunked text from Nepali school textbooks. Schema id — unique segment id source — book/source string book_id — filename-derived id (if present) subject — subject label (e.g., "math") grade — class/grade chapter_index — chapter number chapter_title — chapter/unit name segment_index — index within chapter text — content chunk tokens_approx — rough token… See the full description on the dataset page: https://huggingface.co/datasets/dineshkarki/nepali-textbooks-math-grade10.

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes13downloads
settings

This repository belongs to dineshkarki on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namenepali-textbooks-math-grade10
visibilitypublic
licencenot set
gatedno
ownerdineshkarki
Account settings
dineshkarki/nepali-textbooks-math-grade10 · CoolFace