atomkwk/srt-SpokenCantoneseToWrittenChinese
#Introduction to this dataset This data set is for training llm to translate spoken cantonese srt to written chinese srt(Words in this set is written as simplified chinese characters). Each input and output contain a group of 10 sentances, with a next line character \n between each sentance.
01
#Introduction to this dataset
This data set is for training llm to translate spoken cantonese srt to written chinese srt(Words in this set is written as simplified chinese characters). Each input and output contain a group of 10 sentances, with a next line character \n between each sentance.
