CoolFace
Datasetpublic

atomkwk/srt-SpokenCantoneseToWrittenChinese

#Introduction to this dataset This data set is for training llm to translate spoken cantonese srt to written chinese srt(Words in this set is written as simplified chinese characters). Each input and output contain a group of 10 sentances, with a next line character \n between each sentance.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes1downloads
6 commits on main
4e8cfab2y ago

Upload srt_trainingDataset.json

atomkwk
5aa38692y ago

Update README.md

atomkwk
ed8df542y ago

Update README.md

atomkwk
5487b5b2y ago

Update README.md

atomkwk
461b4172y ago

Test

atomkwk
5c4c29a2y ago

initial commit

atomkwk