vladak/string_ppi_human_1M
STRING PPI Human 1M This dataset contains 1 million human protein–protein interactions (PPIs) derived from STRING v11.5. Columns: seq_a, seq_b: Amino acid sequences of the interacting proteins (≤2048 AA). seq_name_a, seq_name_b: Protein names from STRING. score: Combined score from STRING (0–1000, normalized to 0–1). This score integrates various evidence channels (experimental data, text mining, co-expression, etc.) into a single confidence metric. label: Binary interaction… See the full description on the dataset page: https://huggingface.co/datasets/vladak/string_ppi_human_1M.
015
