ArtificialAnalysis/AA-LCR
Artificial Analysis Long Context Reasoning (AA-LCR) Dataset AA-LCR includes 100 hard text-based questions that require reasoning across multiple real-world documents, with each document set averaging ~100k input tokens. Questions are designed such that answers cannot be directly retrieved from documents and must instead be reasoned from multiple information sources. New in Version 1.1 (September 2026) Sixteen corrected answer keys. Each one was re-verified… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/AA-LCR.
v1.1: correct 16 answer keys and document the judge system prompt (#12)
add document ordering snippets (#10)
fix: remedy question string encoding issue
Update README.md
Update README.md
fix: line breaks
feat: add license notes, format README
Update README.md
Upload AA-LCR_extracted-text.zip
Update README.md
fix: update AA-LCR_dataset.csv token counts
Upload AA-LCR_Dataset.csv (#5)
Update README.md
Update README.md
Upload AA-LCR_Dataset.csv (#1)
initial commit
