CoolFace
Datasetpublic

jie-jw-wu/HumanEvalComm

HumanEvalComm: Benchmarking the Communication Skills of Code Generation for LLMs and LLM Agent πŸ“„ Paper β€’ πŸ’» GitHub Repository β€’ πŸ€— Dataset Viewer Dataset Description HumanEvalComm is a benchmark dataset for evaluating the communication skills of Large Language Models (LLMs) in code generation tasks. It is built upon the widely used HumanEval benchmark. HumanEvalComm contains 762 modified problem descriptions based on the 164 problems in… See the full description on the dataset page: https://huggingface.co/datasets/jie-jw-wu/HumanEvalComm.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes177downloads
settings

This repository belongs to jie-jw-wu on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameHumanEvalComm
visibilitypublic
licenceapache-2.0
gatedno
ownerjie-jw-wu
Account settings
jie-jw-wu/HumanEvalComm Β· CoolFace