CoolFace
Datasetpublic

kowo-co/babble-corrections

babble — corrections Training data for babble: a ~3M parameter byte-level transformer that started from random weights and has only ever learned from people correcting it in Discord. There is no pretraining corpus. There is no scraped chat history. Every row here is somebody deliberately teaching a small confused model to talk. How a row happens Someone @mentions the bot. The bot replies with whatever its current weights produce. Early on this is noise, and it is… See the full description on the dataset page: https://huggingface.co/datasets/kowo-co/babble-corrections.

sourceHugging Facemitupdated 1mo agoView on Hugging Face
0likes28downloads

kowo-co/babble-corrections · main · files are served by the source, never re-hosted here