CoolFace
Datasetpublic

CodeSoulco/TextInsightBench

TextInsightBench English | 简体中文 A natural-language data-mining benchmark for agents: 50 tasks, 435,000 task documents and 944,468 unlabeled learning documents. Each task provides 5,000 or 10,000 texts and a research objective. Agents choose the patterns, populations and comparisons to investigate, then submit up to three findings with complete document assignments, exact quotations, statistics, counterexamples and limitations. Any analysis method is allowed. Contents… See the full description on the dataset page: https://huggingface.co/datasets/CodeSoulco/TextInsightBench.

sourceHugging Faceotherupdated 4d agoView on Hugging Face
0likes970downloads
9 commits on main
f2f99df4d ago

Remove development experiment results

erwinmsmith
a1a5c149d ago

Sync benchmark documentation, evaluation protocol and results (part 2)

erwinmsmith
e8dd8ed9d ago

Sync benchmark documentation, evaluation protocol and results

erwinmsmith
149da1711d ago

Rebuild open exploration tasks and corpus-grounded evaluation

erwinmsmith
e7bfffe11d ago

Document public reference access and the hard mining profile

erwinmsmith
5fda14411d ago

Clarify reference documentation without maintainer attribution

erwinmsmith
b37998c11d ago

Use the TextInsightBench name and descriptive task identifiers

erwinmsmith
7135d1912d ago

Release TextInsightBench v5.1 research dataset

erwinmsmith
87b7f9a12d ago

initial commit

erwinmsmith