CoolFace
Datasetpublic

wordsum/for-the-small-shield-instruct

For The Small Shield — Instruction Data The training data to fine-tune an LLM is derived from a 1.2-million-word manuscript called For The Small Shield (https://github.com/wordsum/For_The_Small_Shield), which I open-sourced 9 years ago. For The Small Shield is grimdark, so the QA pairs may be grimdark. The system role in the training files contains the only words I wrote in the dataset and are intended to make the model just darkish. I've used this to fine-tune a Llama model… See the full description on the dataset page: https://huggingface.co/datasets/wordsum/for-the-small-shield-instruct.

sourceHugging Facecc-by-nc-4.0updated 2mo agoView on Hugging Face
0likes14downloads
5 commits on main
5042db22mo ago

Write to help define and tell the words are grimdark. Write to tell the fact I did write something in the training set. ...Edit model/dataset where needed. Publish for change.

KalabOster
2ccdc472mo ago

Write and edit for clarity and focus definition on the data.

KalabOster
5c2649d2mo ago

Publish the words I wrote to the file.

KalabOster
3e888b82mo ago

Upload folder using huggingface_hub

KalabOster
d2b91cb2mo ago

initial commit

KalabOster