CoolFace
Datasetpublic

viktor-shcherb/longbench2-128k-plus

LongBench2-128k-plus LongBench2-128k-plus is a long-context corpus derived from the zai-org/LongBench-v2 benchmark. It keeps only the "long" examples and exposes just the raw long documents, making it convenient for: long-context pretraining or continued training, long-context adaptation (e.g., RoPE scaling, attention tuning), retrieval and RAG-style experimentation where only documents are needed. All question/answer and multiple-choice metadata from LongBench v2 are dropped;… See the full description on the dataset page: https://huggingface.co/datasets/viktor-shcherb/longbench2-128k-plus.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
0likes15downloads
settings

This repository belongs to viktor-shcherb on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namelongbench2-128k-plus
visibilitypublic
licenceapache-2.0
gatedno
ownerviktor-shcherb
Account settings
viktor-shcherb/longbench2-128k-plus · CoolFace