CoolFace
Datasetpublicgated

IMA-Taiwan/taigi-literature-achiak

Dataset Summary The dataset contains 609 rows. These paragraphs are extracted from authorized prose written by Liau Tiunn Chiak廖張皭 and contain multiple sentences in Taiwanese Taigi written with Hanji. The dataset maintains the original literary style and structure, making it useful for training language models, natural language processing (NLP), and Taiwanese literature research. Dataset Structure Number of rows: 609 (each representing a paragraph) Features:… See the full description on the dataset page: https://huggingface.co/datasets/IMA-Taiwan/taigi-literature-achiak.

sourceHugging Faceccupdated 1y agoView on Hugging Face
0likes8downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.