Venkatdatta/fol-data
FOL Reasoning Dataset A preprocessed and vocabulary-augmented dataset derived from the ProofWriter (Kaggle) OWA splits, built for training a Natural Language → First-Order Logic translation model. The source dataset contains natural-language premises and questions in English along with structured proof metadata. Our preprocessing adds two things that the original does not provide: FOL translations — each natural-language statement is converted to First-Order Logic via a… See the full description on the dataset page: https://huggingface.co/datasets/Venkatdatta/fol-data.
Update dataset card: fix source attribution, premises field, qdep range, ProofWriter vocab sizes, license, citations
Fix premises field, qdep range, license (CC BY-NC-SA 4.0), remove Allen AI reference
Fix source attribution (Kaggle), clarify FOL translation as key preprocessing step, remove script reference
Fix task categories in dataset card
Add dataset card
Upload test.jsonl with huggingface_hub
Upload dev.jsonl with huggingface_hub
Upload train.jsonl with huggingface_hub
initial commit
