assay
Datasets
All datasets matching “assay”assaybench
AssayBench: An Assay-Level Virtual Cell Benchmark
Dataset Summary
AssayBench is a benchmark for evaluating computational models on phenotypic CRISPR screen prediction — a core capability of the "virtual cell" paradigm. It contains 1,901 curated CRISPR screen entries derived from 1,565 unique screens in BioGRID ORCS (version 2025), spanning five major classes of cellular phenotypes. Given a textual description of a CRISPR screen experiment, the task is to predict a… See the full description on the dataset page: https://huggingface.co/datasets/Genentech/assaybench.assay-transfer-record-level-v27-bbb-martins-l3-intern
BBB Martins record-level V27 L3
V27 uses hash-pinned latest V10 evidence and parent-only Gold-v1 evaluation
cohorts. Targeted 8-75-record buckets are held out for OOD evaluation while
minimizing removed training records. ID and OOD queries are capped separately
at 500 in equal bucket rounds. It
renders verified parent SMILES and canonical measurement/unit pairs with atomic
source fallback. Training is balanced before parent-Morgan ranking. L5 is
intentionally excluded.
Rows:… See the full description on the dataset page: https://huggingface.co/datasets/jiosephlee/assay-transfer-record-level-v27-bbb-martins-l3-intern.assay-transfer-record-level-v27-bbb-martins-l2-intern
BBB Martins record-level V27 L2
V27 uses hash-pinned latest V10 evidence and parent-only Gold-v1 evaluation
cohorts. Targeted 8-75-record buckets are held out for OOD evaluation while
minimizing removed training records. ID and OOD queries are capped separately
at 500 in equal bucket rounds. It
renders verified parent SMILES and canonical measurement/unit pairs with atomic
source fallback. Training is balanced before parent-Morgan ranking. L5 is
intentionally excluded.
Rows:… See the full description on the dataset page: https://huggingface.co/datasets/jiosephlee/assay-transfer-record-level-v27-bbb-martins-l2-intern.assay-transfer-record-level-v27-bbb-martins-l4-intern
BBB Martins record-level V27 L4
V27 uses hash-pinned latest V10 evidence and parent-only Gold-v1 evaluation
cohorts. Targeted 8-75-record buckets are held out for OOD evaluation while
minimizing removed training records. ID and OOD queries are capped separately
at 500 in equal bucket rounds. It
renders verified parent SMILES and canonical measurement/unit pairs with atomic
source fallback. Training is balanced before parent-Morgan ranking. L5 is
intentionally excluded.
Rows:… See the full description on the dataset page: https://huggingface.co/datasets/jiosephlee/assay-transfer-record-level-v27-bbb-martins-l4-intern.assay-transfer-record-level-v27-bbb-martins-l1-intern
BBB Martins record-level V27 L1
V27 uses hash-pinned latest V10 evidence and parent-only Gold-v1 evaluation
cohorts. Targeted 8-75-record buckets are held out for OOD evaluation while
minimizing removed training records. ID and OOD queries are capped separately
at 500 in equal bucket rounds. It
renders verified parent SMILES and canonical measurement/unit pairs with atomic
source fallback. Training is balanced before parent-Morgan ranking. L5 is
intentionally excluded.
Rows:… See the full description on the dataset page: https://huggingface.co/datasets/jiosephlee/assay-transfer-record-level-v27-bbb-martins-l1-intern.assay-transfer-record-level-v27-bioavailability-ma-l2-intern
Oral bioavailability record-level V27 L2
V27 uses hash-pinned latest V10 evidence and parent-only Gold-v1 evaluation
cohorts. Targeted 8-75-record buckets are held out for OOD evaluation while
minimizing removed training records. ID and OOD queries are capped separately
at 500 in equal bucket rounds. It
renders verified parent SMILES and canonical measurement/unit pairs with atomic
source fallback. Training is balanced before parent-Morgan ranking. L5 is
intentionally excluded.… See the full description on the dataset page: https://huggingface.co/datasets/jiosephlee/assay-transfer-record-level-v27-bioavailability-ma-l2-intern.
