IMA-Taiwan/taigi-literature-ngkh
Dataset Summary The dataset contains 980 rows. These paragraphs are extracted from authorized paper written by Ngoo Ka Hun吳嘉芬 and contain multiple sentences in Taiwanese Taigi written with Hanji. The dataset maintains the original literary style and structure, making it useful for training language models, natural language processing (NLP), and Taiwanese literature research. Dataset Structure Number of rows: 980 (each representing a paragraph) Features: title:… See the full description on the dataset page: https://huggingface.co/datasets/IMA-Taiwan/taigi-literature-ngkh.
010
No card is published for this repository, or it could not be fetched from Hugging Face right now.
