creativity
openreview_raw
OpenReview Raw
Raw peer review data from OpenReview, covering major ML/AI venues (ICLR, NeurIPS, EMNLP, COLM, ACM MM, and more). Includes reviews, official comments, meta-reviews, and decisions for 49,023 unique papers.
Originally from sumukshashidhar-archive/openreview_raw.
This dataset is a compilation of publicly available data from OpenReview. All original content and data rights belong to OpenReview. This compilation is made available under the Open Data Commons Attribution… See the full description on the dataset page: https://huggingface.co/datasets/creativityschapiro/openreview_raw.ttcw_creativity_eval
Dataset Summary
Stories and annotations for administering the Torrance Test for Creative Writing (TTCW)
Each row in the dataset refers to one specific story, with each column representing the annotations for that specific TTCW category. Each cell contains the story info, information about the TTCW category including the prompt and annotations from 3 different experts on administering the specific TTCW test for the given story.
More info
Repo:… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/ttcw_creativity_eval.details_Undi95__CreativityEngine
Dataset Card for Evaluation run of Undi95/CreativityEngine
Dataset Summary
Dataset automatically created during the evaluation run of model Undi95/CreativityEngine on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Undi95__CreativityEngine.creativity"The only difference between Science and screwing around is writing it down." (Adam Savage)
The LLM Creativity benchmark
Last benchmark update: 28 May 2024
The goal of this benchmark is to evaluate the ability of Large Language Models to be used
as an uncensored creative writing assistant. Human evaluation of the results is done manually,
by me, to assess the quality of writing.
There are 24 questions, some standalone, other follow-ups to previous questions for a multi-turn… See the full description on the dataset page: https://huggingface.co/datasets/froggeric/creativity.openreview-reviews-ft-dataCreativityBench
Dataset Card for CreativityBench
Dataset Details
Dataset Description
CreativityBench is a benchmark for evaluating creative reasoning through affordance-based tool repurposing. Each example places a model or agent in a grounded household scenario and asks it to solve a practical problem by identifying a plausible object part and using that part's annotated affordances.
Curated by: Cheng Qian, Hyeonjeong Ha, Jiayu Liu, Jeonghwan Kim, Jiateng Liu… See the full description on the dataset page: https://huggingface.co/datasets/chengq9/CreativityBench.
