wordsum/for-the-small-shield-instruct
For The Small Shield — Instruction Data The training data to fine-tune an LLM is derived from a 1.2-million-word manuscript called For The Small Shield (https://github.com/wordsum/For_The_Small_Shield), which I open-sourced 9 years ago. For The Small Shield is grimdark, so the QA pairs may be grimdark. The system role in the training files contains the only words I wrote in the dataset and are intended to make the model just darkish. I've used this to fine-tune a Llama model… See the full description on the dataset page: https://huggingface.co/datasets/wordsum/for-the-small-shield-instruct.
Write to help define and tell the words are grimdark. Write to tell the fact I did write something in the training set. ...Edit model/dataset where needed. Publish for change.
Write and edit for clarity and focus definition on the data.
Publish the words I wrote to the file.
Upload folder using huggingface_hub
initial commit
