CoolFace
Datasetpublic

picard47at/punctuation_restoration_900_complex

# punctuation_restoration ## Dataset Summary This dataset is designed for **instruction fine-tuning** of large language models (LLMs), especially for the **Qwen3** family, to perform **punctuation restoration** on Mandarin Chinese text. It is derived from the [`AWeirdDev/zh-tw-articles-6k`](https://huggingface.co/datasets/AWeirdDev/zh-tw-articles-6k) dataset. The `context` field is processed to create input-output pairs in the Qwen3-style message format. - 🔧 **User message**: A cleaned… See the full description on the dataset page: https://huggingface.co/datasets/picard47at/punctuation_restoration_900_complex.

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes20downloads
settings

This repository belongs to picard47at on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namepunctuation_restoration_900_complex
visibilitypublic
licencenot set
gatedno
ownerpicard47at
Account settings
picard47at/punctuation_restoration_900_complex · CoolFace