CoolFace
Datasetpublic

Kausp11/searchlm-eval-caches

searchlm-eval-caches Evaluation caches used by every doc-PPL / dynamic / QA number in RESULTS.md: v2rL_{val,train}_ob_cache.json (reservoir-sampled holdout/train docs + gold blocks), v2rL_val_shuf_cache.json (shuffled-content control), ood_cache.json + ood_fineweb*.jsonl (FineWeb OOD), rank/query caches, items.jsonl. Load with evals_kf/*.py from Search-LLM analysis/. Sources on RCP: /mloscratch/homes/ponkshe/search-lm/evals_kf/data Pushed 2026-09-15 by… See the full description on the dataset page: https://huggingface.co/datasets/Kausp11/searchlm-eval-caches.

sourceHugging Facecc-by-4.0updated 6d agoView on Hugging Face
0likes565downloads
Dataset Card

searchlm-eval-caches

Evaluation caches used by every doc-PPL / dynamic / QA number in RESULTS.md: v2rL{val,train}obcache.json (reservoir-sampled holdout/train docs + gold blocks), v2rLvalshufcache.json (shuffled-content control), oodcache.json + oodfineweb.jsonl (FineWeb OOD), rank/query caches, items.jsonl. Load with evals_kf/.py from Search-LLM analysis/.

Sources on RCP: /mloscratch/homes/ponkshe/search-lm/evals_kf/data

Pushed 2026-09-15 by ops/pushdatasetstohf.py. Provenance: Search-LLM EXPERIMENTPLAN.md / RESULTS.md.