DatarrX/tipitaka-dataset
Myanmar Tipitaka Dataset (Pali Corpus) Dataset Summary The Myanmar Tipitaka Dataset is a high-quality, structured collection of the Buddhist Pali Canon, transcribed in the Myanmar (Burmese) script. This dataset contains 462,504 paragraphs, covering the entire "Triple Basket" (Tipitaka) of Theravada Buddhism, including the original Mula (Canonical texts), Atthakatha (Commentaries), and Tika (Sub-commentaries). This project was initiated by DatarrX to provide a… See the full description on the dataset page: https://huggingface.co/datasets/DatarrX/tipitaka-dataset.
This repository belongs to DatarrX on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
