CoolFace
Datasetpublic

maghzal/PathEval

PathEval: A Benchmark for Evaluating Vision-Language Models as Evaluators for Path Planning Overview Despite their promise to perform complex reasoning, large language models (LLMs) have been shown to have limited effectiveness in end-to-end planning. This has inspired an intriguing question: if these models cannot plan well, can they still contribute to the planning framework as a helpful plan evaluator? In this work, we generalize this question to consider LLMs… See the full description on the dataset page: https://huggingface.co/datasets/maghzal/PathEval.

sourceHugging Facemitupdated 2y agoView on Hugging Face
0likes3.1kdownloads
settings

This repository belongs to maghzal on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namePathEval
visibilitypublic
licencemit
gatedno
ownermaghzal
Account settings
maghzal/PathEval · CoolFace