rationales
vinvl-base-finetuned-hl-rationales-image-captioningqwen-chess-0.5b-rationalesblip-base-captioning-ft-hl-rationalesgit-base-captioning-ft-hl-rationalesautotrain-movie-rationales-1734060527ESM_movie_rationales_defaultT5-small-lora-aqua-rat-gemma-rationales-400-samplesT5-small-lora-aqua-rat-gemma-rationales-1000-samples
movie_rationalesThe movie rationale dataset contains human annotated rationales for movie
reviews.LitBench-RationalesIf you are the author of any comment in this dataset and would like it removed, please contact us and we will comply promptly.
fair-rationalesExplainability methods are used to benchmark
the extent to which model predictions align
with human rationales i.e., are 'right for the
right reasons'. Previous work has failed to acknowledge, however,
that what counts as a rationale is sometimes subjective. This paper
presents what we think is a first of its kind, a
collection of human rationale annotations augmented with the annotators demographic information.SFT_PN_Rationales
SFT Dataset
generated from Qwen/Qwen3-VL-32B-Instruct
verified from OpenGVLab/InternVL3-78B
Domain Distribution of Positive/Negative Rationales
Per-Dataset Positive/Negative Rationale Counts by Domain
litbench-rationales-gpt4
LitBench Rationales - GPT-4 Rubric Evaluations
This dataset contains new rationales for story pair evaluations from the LitBench dataset, generated using GPT-4 with a structured rubric-based evaluation approach.
Evaluation Rubric
The rationales were generated using a 5-criterion rubric:
Creativity & Originality (25 points): Uniqueness of concept, innovative elements, fresh perspective
Writing Quality & Style (25 points): Prose quality, voice consistency, grammar and… See the full description on the dataset page: https://huggingface.co/datasets/SAA-Lab/litbench-rationales-gpt4.LitBench-new-rationales
Dataset Card for "LitBench-Rationales-GPT4-Complete"
More Information needed
