CoolFace
Datasetpublicgated

IMA-Taiwan/zhtw-literature-ots

Dataset Summary The dataset contains 2,349 rows. These paragraphs are extracted from authorized novels written by Ou Tiong Siong胡長松 and contain multiple sentences in Traditional Chinese. The dataset maintains the original literary style and structure, making it useful for training language models, natural language processing (NLP), and Taiwanese literature research. Dataset Structure Number of rows: 2,349 (each representing a paragraph) Features: title: Book… See the full description on the dataset page: https://huggingface.co/datasets/IMA-Taiwan/zhtw-literature-ots.

sourceHugging Faceccupdated 1y agoView on Hugging Face
1likes11downloads

No commit history came back for main. The revision may not exist, or the source declined the request.