CoolFace
Datasetpublic

viktor-shcherb/longbench2-128k-plus

LongBench2-128k-plus LongBench2-128k-plus is a long-context corpus derived from the zai-org/LongBench-v2 benchmark. It keeps only the "long" examples and exposes just the raw long documents, making it convenient for: long-context pretraining or continued training, long-context adaptation (e.g., RoPE scaling, attention tuning), retrieval and RAG-style experimentation where only documents are needed. All question/answer and multiple-choice metadata from LongBench v2 are dropped;… See the full description on the dataset page: https://huggingface.co/datasets/viktor-shcherb/longbench2-128k-plus.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
0likes15downloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

viktor-shcherb/longbench2-128k-plus · main · files are served by the source, never re-hosted here