CoolFace
Datasetpublic

sabin1234/NEPALI-MCQ-SFT-MULTIDOMAIN-DATASET

Nepali Devanagari SFT Dataset — Final Clean Release A 100,000-row synthetic Nepali SFT dataset designed for Nepali-language instruction-following and supervised fine-tuning experiments. Release status: Final structural and Unicode validation passed for the previously identified contamination/corruption patterns. Dataset at a Glance Property Value Total rows 100,000 Total conversation messages 200,000 Human messages 100,000 GPT messages 100,000… See the full description on the dataset page: https://huggingface.co/datasets/sabin1234/NEPALI-MCQ-SFT-MULTIDOMAIN-DATASET.

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
0likes20downloads
14 commits on main
71df8482mo ago

Update README.md

sabin1234
d8ee4822mo ago

Upload README.md with huggingface_hub

sabin1234
25bd86f2mo ago

Delete README.md

sabin1234
9830a202mo ago

Create README.md

sabin1234
1b7c7a82mo ago

Update README_all_pure_nepali_FINAL_CLEAN_FINAL(1).md

sabin1234
38ab5cf2mo ago

Delete README.md

sabin1234
c0d15e52mo ago

Upload README_all_pure_nepali_FINAL_CLEAN_FINAL(1).md

sabin1234
43f58aa2mo ago

Update README.md

sabin1234
842af4e2mo ago

Upload all_pure_nepali_FINAL_CLEAN_FINAL.json

sabin1234
a2c5be12mo ago

Delete all_pure_nepali_FINAL_CLEAN_FINAL_recheck.json

sabin1234
45fd2c62mo ago

Update README.md

sabin1234
1e5effd2mo ago

Upload README.md

sabin1234
0f637022mo ago

Upload all_pure_nepali_FINAL_CLEAN_FINAL_recheck.json

sabin1234
63a2dad2mo ago

initial commit

sabin1234