CodeSoulco/TextInsightBench
TextInsightBench English | 简体中文 A natural-language data-mining benchmark for agents: 50 tasks, 435,000 task documents and 944,468 unlabeled learning documents. Each task provides 5,000 or 10,000 texts and a research objective. Agents choose the patterns, populations and comparisons to investigate, then submit up to three findings with complete document assignments, exact quotations, statistics, counterexamples and limitations. Any analysis method is allowed. Contents… See the full description on the dataset page: https://huggingface.co/datasets/CodeSoulco/TextInsightBench.
Remove development experiment results
Sync benchmark documentation, evaluation protocol and results (part 2)
Sync benchmark documentation, evaluation protocol and results
Rebuild open exploration tasks and corpus-grounded evaluation
Document public reference access and the hard mining profile
Clarify reference documentation without maintainer attribution
Use the TextInsightBench name and descriptive task identifiers
Release TextInsightBench v5.1 research dataset
initial commit
