CoolFace
Datasetpublic

paodigitalhub/pao-sentences-dataset

Pa'O Sentences Dataset is a text corpus for Pa'O language (ပအိုဝ်ႏ) containing structured line-by-line sentences designed for NLP, LLM pre-training, and machine translation. 📝 Pa'O Sentences Dataset (ပအိုဝ်ႏ လိက်လာႏငေါဝ်းရဲဉ်ႏ ရွမ်ခြွဉ်းဗူႏ) 📌 Project Summary (ထာꩻမာꩻခြပ်ရဲဉ်ႏ နပ်ထွားရဲပ်အအဲဉ်ႏ) The Pa'O Sentences Dataset is an open-source textual corpus developed to support Natural Language Processing (NLP), Large Language Model (LLM) pre-training, Machine… See the full description on the dataset page: https://huggingface.co/datasets/paodigitalhub/pao-sentences-dataset.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
1likes84downloads
settings

This repository belongs to paodigitalhub on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namepao-sentences-dataset
visibilitypublic
licencecc-by-4.0
gatedno
ownerpaodigitalhub
Account settings
paodigitalhub/pao-sentences-dataset · CoolFace