CoolFace
Datasetpublic

opencompass/NeedleBench

Dataset Description Dataset Summary The NeedleBench dataset is a part of the OpenCompass project, designed to evaluate the capabilities of large language models (LLMs) in processing and understanding long documents. It includes a series of test scenarios that assess models' abilities in long text information extraction and reasoning. The dataset is structured to support tasks such as single-needle retrieval, multi-needle retrieval, multi-needle reasoning, and… See the full description on the dataset page: https://huggingface.co/datasets/opencompass/NeedleBench.

sourceHugging Facemitupdated 4mo agoView on Hugging Face
6likes8.9kdownloads
settings

This repository belongs to opencompass on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameNeedleBench
visibilitypublic
licencemit
gatedno
owneropencompass
Account settings
opencompass/NeedleBench · CoolFace