CoolFace
Datasetpublic

jie-jw-wu/HumanEvalComm

HumanEvalComm: Benchmarking the Communication Skills of Code Generation for LLMs and LLM Agent 📄 Paper • 💻 GitHub Repository • 🤗 Dataset Viewer Dataset Description HumanEvalComm is a benchmark dataset for evaluating the communication skills of Large Language Models (LLMs) in code generation tasks. It is built upon the widely used HumanEval benchmark. HumanEvalComm contains 762 modified problem descriptions based on the 164 problems in… See the full description on the dataset page: https://huggingface.co/datasets/jie-jw-wu/HumanEvalComm.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes177downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
jie-jw-wu/HumanEvalComm · CoolFace