datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ShortStory-SFT-jsonl
Public Domain Short Fiction with Prompts
719 complete short stories (300 to 2500 words) by 26 authors whose work is in the public domain,
each paired with a natural-language request that could plausibly have produced it. Built for supervised
fine-tuning of small language models on fiction, where the usual sources (forum stories, model-generated
stories) lack the structural control of published short fiction.
Fields
field
description
id
stable id (hash… See the full description on the dataset page: https://huggingface.co/datasets/Travis-ML/ShortStory-SFT-jsonl.shortstories_synthlabelsWIP
Nothing to see here, it's just some shortstories from internet with synthetic writing prompts
