jet-ai/pg19-subsample
Jet-Long Evaluation Datasets This repository contains the evaluation data used in the paper Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE. Code: GitHub Repository Dataset Description These datasets are processed versions of standard benchmarks used to evaluate the long-context capabilities of Jet-Long: RULER-500: A dataset for evaluating long-context understanding and retrieval up to 128K context. PG-19-subsample: A subsampled version of… See the full description on the dataset page: https://huggingface.co/datasets/jet-ai/pg19-subsample.
024
