datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
proofwriter-deduction-balancedA processed subset of the OWA section of the ProofWriter dataset.
Each train/test split contains 300 entries, each of which has a unique set of theories and a single question for those theories.
Both splits are balanced so that the depth of the proof required to answer the question varies evenly between 0-5 (50 entries each), and the labels are balanced (100 each).
'Unknown' labels have been replaced by 'Uncertain' to match other datasets.
ProofWriter
Github
https://github.com/teacherpeterpan/Logic-LLM/blob/main/outputs/logic_programs/ProofWriter_dev_gpt-4.json
Reference
@inproceedings{PanLogicLM23,
author = {Liangming Pan and
Alon Albalak and
Xinyi Wang and
William Yang Wang},
title = {{Logic-LM:} Empowering Large Language Models with Symbolic Solvers for Faithful Logical Reasoning},
booktitle = {Findings of the 2023 Conference on Empirical… See the full description on the dataset page: https://huggingface.co/datasets/renma/ProofWriter.proofwriter-datasetproofwriter-datasetstream-2-proofwriterproofwriter-deduction-balancedA processed subset of the OWA section of the ProofWriter dataset.
Each train/test split contains 300 entries, each of which has a unique set of theories and a single question for those theories.
Both splits are balanced so that the depth of the proof required to answer the question varies evenly between 0-5 (50 entries each), and the labels are balanced (100 each).
'Unknown' labels have been replaced by 'Uncertain' to match other datasets.
