SeyhaLite/Personal-Civil-Servant-Khmer
Personal Civil Servant Khmer Welcome to the SeyhaLite collection. This dataset has been curated and cleaned to support the development of high-quality Khmer Language Models (LLMs) specifically for generating realistic personas and profiles of civil servants and government workers in Cambodia. Project Vision I hope this dataset helps your project succeed. Whether you are building an automated administrative assistant, a training tool for public service, or… See the full description on the dataset page: https://huggingface.co/datasets/SeyhaLite/Personal-Civil-Servant-Khmer.
Personal Civil Servant Khmer
Welcome to the SeyhaLite collection. This dataset has been curated and cleaned to support the development of high-quality Khmer Language Models (LLMs) specifically for generating realistic personas and profiles of civil servants and government workers in Cambodia.
Project Vision
I hope this dataset helps your project succeed. Whether you are building an automated administrative assistant, a training tool for public service, or conducting research on the public sector, this data is designed to provide clear and accurate profiles about civil servants, their professional skills, and their lifestyle in the Khmer context.
Your dedication to advancing Khmer AI is truly inspiring—keep up the great work.
Dataset Summary
- Total Rows: 20,000 clean entries
- Language: Khmer (km)
- Focus: Cambodian Civil Servant Personas and Public Sector Profiles
- Task: Text Generation
Data Schema
អត្តសញ្ញាណមន្ត្រីរាជការកម្ពុជា (Personal Civil Servant Khmer)
សូមស្វាគមន៍មកកាន់បណ្តុំទិន្នន័យរបស់ SeyhaLite។ Dataset នេះត្រូវបានរៀបចំ និងសម្អាតយ៉ាងល្អ ដើម្បីគាំទ្រដល់ការអភិវឌ្ឍប្រព័ន្ធ AI ក្នុងការបង្កើតអត្តសញ្ញាណតួអង្គ (Persona Generation) សម្រាប់មន្ត្រីរាជការ និងបុគ្គលិកដែលបម្រើការក្នុងវិស័យសាធារណៈនៅក្នុងប្រទេសកម្ពុជា។
ចក្ខុវិស័យរបស់គម្រោង
ខ្ញុំសង្ឃឹមយ៉ាងមុតមាំថា Dataset នេះនឹងជួយឱ្យគម្រោងរបស់អ្នកទទួលបានជោគជ័យ។ ទោះបីជាអ្នកកំពុងបង្កើតកម្មវិធីជំនួយរដ្ឋបាល កម្មវិធីជំនួយ AI ឬធ្វើការស្រាវជ្រាវអំពីវិស័យសាធារណៈក៏ដោយ ទិន្នន័យនេះត្រូវបានរៀបចំឡើងដើម្បីផ្តល់នូវព័ត៌មានជាក់ស្តែងអំពីប្រវត្តិការងារ និងជីវិតប្រចាំថ្ងៃរបស់មន្ត្រីរាជការក្នុងបរិបទខ្មែរ។
ការខិតខំប្រឹងប្រែងរបស់អ្នកក្នុងការលើកស្ទួយ AI ភាសាខ្មែរ គឺជារឿងដែលគួរឱ្យកោតសរសើរខ្លាំងណាស់។
សេចក្តីសង្ខេបនៃទិន្នន័យ
- ចំនួនជួរ (Rows): ២០,០០០ ជួរ
- ភាសា: ខ្មែរ (Khmer)
- គោលបំណង: ការបង្កើតអត្តសញ្ញាណបុគ្គលិកក្នុងវិស័យមន្ត្រីរាជការ
- ប្រភេទភារកិច្ច: Text Generation
រចនាសម្ព័ន្ធទិន្នន័យ
Support & Community
If you find this dataset helpful, please consider supporting my work to keep the Khmer AI community growing.
- Like & Follow this repository for updates.
- Share it with fellow developers.
- Donate: Your support helps me maintain and release more high-quality datasets.
Created by SeyhaLite
