CoolFace
Datasetpublic

ris3abh-11/engram-eval

engram evaluation data The evaluation data behind Typed Decisions in Agent Memory: Where They Help, Where They Don't, and What It Costs (Rishabh Sharma, 2026, doi:10.5281/zenodo.22948964; version 1: doi:10.5281/zenodo.22941758): update sets that extend LoCoMo with fact changes, labeled contradiction pairs, the relation decisions escalated to an LLM, and every scored answer from the paper's runs. Code: the engram repository (bench/make_hf_dataset.py builds this directory from the… See the full description on the dataset page: https://huggingface.co/datasets/ris3abh-11/engram-eval.

sourceHugging Facecc-by-nc-4.0updated 19h agoView on Hugging Face
0likes19downloads

ris3abh-11/engram-eval · main · files are served by the source, never re-hosted here