CoolFace
Datasetpublicgated

IMA-Taiwan/taigi-literature-ngkh

Dataset Summary The dataset contains 980 rows. These paragraphs are extracted from authorized paper written by Ngoo Ka Hun吳嘉芬 and contain multiple sentences in Taiwanese Taigi written with Hanji. The dataset maintains the original literary style and structure, making it useful for training language models, natural language processing (NLP), and Taiwanese literature research. Dataset Structure Number of rows: 980 (each representing a paragraph) Features: title:… See the full description on the dataset page: https://huggingface.co/datasets/IMA-Taiwan/taigi-literature-ngkh.

sourceHugging Faceccupdated 1y agoView on Hugging Face
0likes10downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.