Yiyang-Ian-Li/LongDA
LongDA Dataset Card Dataset Description LongDA is a data analysis benchmark for evaluating LLM-based agents under documentation-intensive analytical workflows. It features authentic U.S. government survey data with complete, long documentation, testing LLMs' ability to navigate complex real-world datasets before performing analysis. Dataset Summary 505 queries extracted from 30 expert-written publications 17 U.S. national surveys covering health… See the full description on the dataset page: https://huggingface.co/datasets/Yiyang-Ian-Li/LongDA.
Fix benchmark metadata typos
Fix benchmark ground truth values
update readme
Move large CSV files to Git LFS
Track large CSV files with Git LFS
update README
Update dataset card
Update dataset card
Move benchmark.csv out of LFS for dataset preview
Upload remaining benchmark files
Track large/binary files with Git LFS
Track .xpt files with Git LFS
Add .xpt via LFS
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
initial commit
