CoolFace
Datasetpublic

opencompass/NeedleBench

Dataset Description Dataset Summary The NeedleBench dataset is a part of the OpenCompass project, designed to evaluate the capabilities of large language models (LLMs) in processing and understanding long documents. It includes a series of test scenarios that assess models' abilities in long text information extraction and reasoning. The dataset is structured to support tasks such as single-needle retrieval, multi-needle retrieval, multi-needle reasoning, and… See the full description on the dataset page: https://huggingface.co/datasets/opencompass/NeedleBench.

sourceHugging Facemitupdated 4mo agoView on Hugging Face
6likes8.9kdownloads

opencompass/NeedleBench · main · files are served by the source, never re-hosted here