CoolFace
14 results

pddl

Babelscape /PDDL2PRM PDDL2PRM: Planning-Based Step-Level Supervision for Process Reward Models PDDL2PRM is a large-scale dataset for training and evaluating Process Reward Models (PRMs) with fine-grained, step-level supervision derived from symbolic planning problems. Unlike many PRM datasets that rely on human annotation, LLM judges, or final-answer correctness, PDDL2PRM uses Planning Domain Definition Language (PDDL) problems to generate structured reasoning trajectories whose intermediate steps… See the full description on the dataset page: https://huggingface.co/datasets/Babelscape/PDDL2PRM.texttext-classification1M<n<10M0 likes42 downloads3mo agoHugging Facebhoy /agentboard-pddl-primerl-checkpoints0 likes35 downloads3mo agoHugging FaceSelf-CriTeach /pddl-planning-data PDDL Planning Data (Self-CriTeach) PDDL-style planning problem–plan pairs used to train and evaluate Self-CriTeach models. Each example is a single planning problem in the Blocksworld family (and three unseen extensions for OOD evaluation), formatted as static predicates + initial dynamic state + ground-truth action sequence. Companion to: Paper: Self-CriTeach: LLM Self-Teaching and Self-Critiquing for Improving Robotic Planning Code: https://github.com/markli1hoshipu/Plan_LLM… See the full description on the dataset page: https://huggingface.co/datasets/Self-CriTeach/pddl-planning-data.texttext-generation10K<n<100K0 likes35 downloads3mo agoHugging FaceMengkangHu /AgentGen-PDDL0 likes28 downloads10mo agoHugging FaceMohamedYoussef97 /toy_example_sim2real_data_with_loaded_kg_no_pddlimagen<1K0 likes8 downloads10mo agoHugging FaceHoshipu /wm_pddl-dataset0 likes4 downloads8mo agoHugging Face