baber/agieval
Dataset Card for AGIEval Dataset Summary AGIEval is a human-centric benchmark specifically designed to evaluate the general abilities of foundation models in tasks pertinent to human cognition and problem-solving. This benchmark is derived from 20 official, public, and high-standard admission and qualification exams intended for general human test-takers, such as general college admission tests (e.g., Chinese College Entrance Exam (Gaokao) and American SAT), law… See the full description on the dataset page: https://huggingface.co/datasets/baber/agieval.
Fix URL (#1)
fixed bug
clean up
mislabeled rows fixed in repo
added description
fixed bug
added sat w/o passage
added 'passage' column to subsets
Update README.md
Update agieval.py
Update agieval.py
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Update agieval.py
Create agieval.py
initial commit
