datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
creativity"The only difference between Science and screwing around is writing it down." (Adam Savage)
The LLM Creativity benchmark
Last benchmark update: 28 May 2024
The goal of this benchmark is to evaluate the ability of Large Language Models to be used
as an uncensored creative writing assistant. Human evaluation of the results is done manually,
by me, to assess the quality of writing.
There are 24 questions, some standalone, other follow-ups to previous questions for a multi-turn… See the full description on the dataset page: https://huggingface.co/datasets/froggeric/creativity.collabllm-multiturn-creativity
