superseded
superseded-standards-mappings
Superseded and deprecated identifier mappings
Canonical, always-current version: https://referencesource.org/superseded-standards-mappings/
Machine-readable: https://referencesource.org/superseded-standards-mappings/data.json — this mirror is a point-in-time copy.
Last verified: 2026-08-04
Stale after: 2027-01-31 (past this date, prefer the canonical copy —
it re-verifies on a cadence this snapshot does not)
Records: 141
Withdrawn standards, deprecated API models and retired… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/superseded-standards-mappings.bdh-cl_ra2-supersededfactprobe-replication-SUPERSEDED-stage1-counts-olmotok-v1
SUPERSEDED - do not use
Renamed 2026-08-25. Use latkes/factprobe-replication-stage1-counts-canonical-v1 instead.
It was counted with a name list missing each entity's canonical Wikidata name for 10,471 of 42,235 entities (24.8%). "Barack Obama" is not in it, and occurs 38,978,811 times in the corpus it claims to count.
It is kept only so earlier numbers can be traced to where they came from. Nothing current should read it.… See the full description on the dataset page: https://huggingface.co/datasets/latkes/factprobe-replication-SUPERSEDED-stage1-counts-olmotok-v1.factprobe-replication-SUPERSEDED-olmomix-counts-v1
SUPERSEDED - do not use
Renamed 2026-08-25. Use latkes/factprobe-replication-stage1-counts-canonical-v1 instead.
It was produced by querying the infini-gram service, which indexes the corpus with the Llama-2 tokenizer, rather than by counting the corpus as OLMo token sequences. It also predates the canonical-name repair.
It is kept only so earlier numbers can be traced to where they came from. Nothing current should read it.
factprobe-replication-olmomix-counts-v1… See the full description on the dataset page: https://huggingface.co/datasets/latkes/factprobe-replication-SUPERSEDED-olmomix-counts-v1.factprobe-replication-SUPERSEDED-stage2-counts-olmotok-v1
SUPERSEDED - do not use
Renamed 2026-08-25. Use latkes/factprobe-replication-stage2-counts-canonical-v1 instead.
It was counted with a name list missing each entity's canonical Wikidata name for 10,471 of 42,235 entities (24.8%).
It is kept only so earlier numbers can be traced to where they came from. Nothing current should read it.
factprobe-replication-stage2-counts-olmotok-v1
Occurrences of every probed entity name in the OLMo-2-7B mid-training corpus (the… See the full description on the dataset page: https://huggingface.co/datasets/latkes/factprobe-replication-SUPERSEDED-stage2-counts-olmotok-v1.quota-relief-superseded-spb9-ckpts
