scbe
Datasets
All datasets matching “scbe”SCBench-preprocessedThis is the preprocessed version of Microsoft SCBench, used by KVzip:
Each data example has a format of {context: str, question: List[str], answers: List[str]}
Each dataset contains only examples whose context token length (measured with the LLaMA3 tokenizer) is less than 125K, fitting within the context limit of LLaMA3 models.
We also provide shortened versions of SCBench, excluding tasks {choice_eng, qa_eng, and vt}, which are difficult to shorten.
The "tiny" tag (e.g., scbench_kv_tiny)… See the full description on the dataset page: https://huggingface.co/datasets/Jang-Hyun/SCBench-preprocessed.scbe-aethermoore-training-data
Status: canonical. Primary public training dataset for SCBE-AETHERMOORE and the most-used repo in this account. Other scbe-* dataset repos are experiment-specific slices.
SCBE-AETHERMOORE Training Dataset
Supervised fine-tuning (SFT) dataset for the SCBE-AETHERMOORE hyperbolic geometry AI safety and governance framework.
Overview
This dataset contains 10,978 training pairs spanning the full SCBE-AETHERMOORE system: 14-layer architecture knowledge, Six Sacred… See the full description on the dataset page: https://huggingface.co/datasets/issdandavis/scbe-aethermoore-training-data.SCBench
SCBench
[Paper]
[Code]
[Project Page]
SCBench (SharedContextBench) is a comprehensive benchmark to evaluate efficient long-context methods in a KV cache-centric perspective, analyzing their performance across the full KV cache lifecycle (generation, compression, retrieval, and loading) in real-world scenarios where context memory (KV cache) is shared and reused across multiple requests.
🎯 Quick Start
Load Data
You can download and load the SCBench data… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/SCBench.scbench-data
scBench Canonical Data Files
This dataset contains the .h5ad data files for the scBench canonical subset (30 evaluations across 5 platforms).
About scBench
scBench is a benchmark for agentic single-cell RNA-seq analysis. It evaluates whether AI agents can solve practical bioinformatics tasks with deterministic grading.
Paper: https://arxiv.org/abs/2602.09063
GitHub: https://github.com/latchbio/scbench
Inspect Evals: https://github.com/UKGovernmentBEIS/inspect_evals… See the full description on the dataset page: https://huggingface.co/datasets/retroam/scbench-data.scbe-chemistry-sft
Status: experimental. Research artifact, not a production candidate. Canonical dataset: scbe-aethermoore-training-data.
SCBE chemistry SFT
Chemistry adapter training data, split by purpose rather than one flat pile:
chemistry_adapter_invariants_v1_{train,eval}.sft.jsonl - conservation and
invariant rows
chemistry_adapter_verification_v1_{train,eval}.sft.jsonl - verification rows
chemistry_gate_repair_v1_{train,eval}.sft.jsonl - gate-repair rows
aligned_foundations_v2_{train… See the full description on the dataset page: https://huggingface.co/datasets/issdandavis/scbe-chemistry-sft.SCBench
SCBench
[Paper]
[Code]
SCBench (SharedContextBench) is a comprehensive benchmark to evaluate efficient long-context methods in a KV cache-centric perspective, analyzing their performance across the full KV cache lifecycle (generation, compression, retrieval, and loading) in real-world scenarios where context memory (KV cache) is shared and reused across multiple requests.
Dataset
SCBench covers 12 diverse tasks that test four key long-context capabilities: string… See the full description on the dataset page: https://huggingface.co/datasets/MInference/SCBench.
