datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
RPToolkit-demo-datasetRPToolkit is a data generation pipeline, part of Augmentoolkit, that generates synthetic RP sessions inspired by input stories. Basically: feed in Lord of the Rings, get out high fantasy adventure RPs.
This dataset, containing over a million trainable tokens across around 1000 RP sessions, is meant to showcase the capabilities of this pipeline.
The input texts used were: a variety of myths and classic stories from Gutenberg; the first few chapters of some miscellaneous webnovels and… See the full description on the dataset page: https://huggingface.co/datasets/Heralax/RPToolkit-demo-dataset.rp-teacher-synth-dporp-teacher-synth-wizard-bixtral-dpoSmall dataset attempting to instruct a model in the usage of system prompts.
Personas and some 'rejected' entries were synthesized by Lambent/braidbird-scribe-7B or a related local model, and the rest was synthesized with WizardLM-2-8x22B.
rp-teacher-synth-wizard-bixtral-sharegptSmall dataset attempting to instruct a model in the usage of system prompts.
Personas were synthesized by Lambent/braidbird-scribe-7B, and the rest was synthesized with WizardLM-2-8x22B.
These are multi-turn, in-character conversations of a specific number of turns.
Text length was not strictly specified beyond setting Wizard's max output length to 1024 tokens.
Having sampled a couple, my estimate is they should mostly fit within 4096 tokens, and certainly within 8192.
rp-teacher-synth-sharegptrp-testHeralax_RPToolkit-demo-dataset
