alpindale/visual-novels
Visual Novel Dataset This dataset contains parsed Visual Novel scripts for training language models. The dataset consists of approximately 60 million tokens of parsed scripts. Dataset Structure The dataset follows a general structure for visual novel scripts: Dialogue lines: Dialogue lines are formatted with the speaker's name followed by a colon, and the dialogue itself enclosed in quotes. For example: John: "Hello, how are you?" Actions and narration: Actions… See the full description on the dataset page: https://huggingface.co/datasets/alpindale/visual-novels.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face