CoolFace
Datasetpublic

m-a-p/AetherCode

AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions Introduction Competitive programming has emerged as a critical benchmark for evaluating the reasoning and coding capabilities of Large Language Models (LLMs). Despite impressive progress on existing benchmarks, we argue that current evaluations overstate model proficiency, masking a substantial gap between LLMs and elite human programmers. This gap… See the full description on the dataset page: https://huggingface.co/datasets/m-a-p/AetherCode.

sourceHugging Facecc-by-4.0updated 1y agoView on Hugging Face
8likes654downloads
19 commits on main
54cf2411y ago

Improve dataset card: Add metadata, update Hugging Face and paper badges (#2)

zhwang01, nielsr
c2117ca1y ago

Update README.md

zhwang01
f2e85721y ago

Update README.md

zhwang01
83e387a1y ago

Update README.md

zhwang01
ab71d971y ago

Update README.md

zhwang01
ddd16681y ago

Create DISCLAIMER

zhwang01
c26a0641y ago

Update README.md

zhwang01
e51fe191y ago

Upload 2 files

zhwang01
ffd85171y ago

Update README.md

zhwang01
aff2fbd1y ago

Update README.md

zhwang01
86b382c1y ago

Create LICENSE

zhwang01
2324e851y ago

Update README.md

zhwang01
92a65541y ago

Update README.md

zhwang01
dc77a641y ago

Upload dataset

zhwang01
498fc291y ago

Upload dataset

zhwang01
608d8e81y ago

Delete README.md

zhwang01
a8c8ebe1y ago

Delete v1_2025

zhwang01
22abc591y ago

Upload dataset

zhwang01
2a8bd081y ago

initial commit

zhwang01