MATH · IN · MODELS

Protein-LM probes recover structure; attention heads align above background

measured in 1 paper

Vig et al. analyze five protein language models (TapeBert, ProtBert, ProtBert-BFD, ProtAlbert, ProtXLNet) with linear single-layer probes and attention-head analysis [vig-etal-2020-bertology-meets-biology-interpreting-attention-in-protein-language-models] Individual attention heads concentrate on 3D contact maps, binding sites, and secondary structure far above background (up to 63.2% of a head's attention on contacts vs a 1.3% background; one head 49% on binding sites, Bonferroni p<0.00001) [vig-etal-2020-bertology-meets-biology-interpreting-attention-in-protein-language-models] Linear probes on output embeddings recover the same structures, quantified via precision@L/5 for contacts, precision@L/20 for binding sites, and F1 for secondary structure [vig-etal-2020-bertology-meets-biology-interpreting-attention-in-protein-language-models]

Context

convergent evidence for the same biological structure from two different geometric measurements (attention-pattern alignment and embedding-level linear probing) on the same pretrained model, foundational protein-attention-geometry precedent predating most later protein-LM interpretability work in this map

Papers

BERTology Meets Biology: Interpreting Attention in Protein Language Models — Vig, Jesse, Madani, Ali, Varshney, Lav R., Xiong, Caiming, Socher, Richard, Rajani, Nazneen Fatema2020 · arXiv:2006.15222