jkhouja/LingOly-TOO
LingOly-TOO (L2) Links 📊 Website and Leaderboard 📎 Paper 🧩 Code Summary LingOly-TOO (L2) is a challenging linguistics reasoning benchmark designed to counteracts answering without reasoning (e.g. by guessing or memorizing answers). Dataset format LingOly-TOO benchmark was created by generating up to 6 obfuscations per problem for 82 problems source from original LingOly benchmark. Dataset contains over 1200 question answer pairs. Some… See the full description on the dataset page: https://huggingface.co/datasets/jkhouja/LingOly-TOO.
Upload test_small.zip
Delete main_results.svg
Upload main_scores.svg
Upload main_results.svg
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Upload main_scores.svg
Update README.md
Upload main_scores.svg
Update README.md
Upload test_small.zip
Update README.md
Upload test_small.zip
Upload test_small.zip
Upload test_small.zip
Upload test_small.zip
Upload test_small.zip
Upload test_small.zip
Upload test_small.zip
Update README.md
Update README.md
Upload main_scores.svg
Update README.md
Update README.md
Upload test_small.zip
Update README.md
Initial README
initial commit
