SeyhaLite/Personal-Generate-GenZ-Students-Khmer
Personas Generate Genz Students Khmer Welcome to the SeyhaLite collection. This dataset has been curated and cleaned to support the development of high-quality Khmer Language Models (LLMs) specifically focused on persona generation and synthetic character creation for the Gen Z demographic in Cambodia. Project Vision I hope this dataset helps your project succeed! Whether you are building a specialized chatbot, an AI assistant, or conducting social research, this… See the full description on the dataset page: https://huggingface.co/datasets/SeyhaLite/Personal-Generate-GenZ-Students-Khmer.
Personas Generate Genz Students Khmer
Welcome to the SeyhaLite collection. This dataset has been curated and cleaned to support the development of high-quality Khmer Language Models (LLMs) specifically focused on persona generation and synthetic character creation for the Gen Z demographic in Cambodia.
Project Vision
I hope this dataset helps your project succeed! Whether you are building a specialized chatbot, an AI assistant, or conducting social research, this data is designed to provide clear and realistic profiles about Gen Z students, their skills, and their lifestyle in the Khmer context.
Your dedication to advancing Khmer AI is truly inspiring—keep up the great work!
Dataset Summary
- Total Rows: 20,000 clean entries
- Language: Khmer (km)
- Focus: Gen Z Student Personas and Profiles
- Task: Text Generation
Data Schema
ការបង្កើតអត្តសញ្ញាណសិស្សជំនាន់ Gen Z កម្ពុជា (Personas Generate Genz Students Khmer)
សូមស្វាគមន៍មកកាន់បណ្តុំទិន្នន័យរបស់ SeyhaLite។ Dataset នេះត្រូវបានរៀបចំ និងសម្អាតយ៉ាងល្អ ដើម្បីគាំទ្រដល់ការអភិវឌ្ឍប្រព័ន្ធ AI ក្នុងការបង្កើតអត្តសញ្ញាណតួអង្គ (Persona Generation) សម្រាប់យុវជនជំនាន់ Gen Z នៅក្នុងប្រទេសកម្ពុជា។
ចក្ខុវិស័យរបស់គម្រោង
ខ្ញុំសង្ឃឹមយ៉ាងមុតមាំថា Dataset នេះនឹងជួយឱ្យគម្រោងរបស់អ្នកទទួលបានជោគជ័យ! ទោះបីជាអ្នកកំពុងបង្កើត Chatbot, កម្មវិធីជំនួយ ឬធ្វើការស្រាវជ្រាវផ្សេងៗ ទិន្នន័យនេះត្រូវបានរៀបចំឡើងដើម្បីផ្តល់នូវព័ត៌មានជាក់ស្តែងអំពីប្រវត្តិ ចំណូលចិត្ត និងជំនាញរបស់សិស្សនិស្សិតខ្មែរ។
ការខិតខំប្រឹងប្រែងរបស់អ្នកក្នុងការលើកស្ទួយ AI ភាសាខ្មែរ គឺជារឿងដែលគួរឱ្យកោតសរសើរខ្លាំងណាស់!
សេចក្តីសង្ខេបនៃទិន្នន័យ
- ចំនួនជួរ (Rows): ២០,០០០ ជួរ
- ភាសា: ខ្មែរ (Khmer)
- គោលបំណង: ការបង្កើតអត្តសញ្ញាណសិស្សនិស្សិត Gen Z
- ប្រភេទភារកិច្ច: Text Generation
រចនាសម្ព័ន្ធទិន្នន័យ
Support & Community
If you find this dataset helpful, please consider supporting my work to keep the Khmer AI community growing.
- Like & Follow this repository for updates.
- Share it with fellow developers.
- Donate: Your support helps me maintain and release more high-quality datasets.
Created by SeyhaLite
