REVIEW 3 cited by
BERTology Meets Biology: Interpreting Attention in Protein Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Transformer architectures have proven to learn useful representations for protein classification and generation tasks. However, these representations present challenges in interpretability. In this work, we demonstrate a set of methods for analyzing protein Transformer models through the lens of attention. We show that attention: (1) captures the folding structure of proteins, connecting amino acids that are far apart in the underlying sequence, but spatially close in the three-dimensional structure, (2) targets binding sites, a key functional component of proteins, and (3) focuses on progressively more complex biophysical properties with increasing layer depth. We find this behavior to be consistent across three Transformer architectures (BERT, ALBERT, XLNet) and two distinct protein datasets. We also present a three-dimensional visualization of the interaction between attention and protein structure. Code for visualization and analysis is available at https://github.com/salesforce/provis.
Forward citations
Cited by 3 Pith papers
-
Directed Evolution of Proteins via Bayesian Optimization in Embedding Space
A Gaussian-process Bayesian optimizer running in protein language model embedding space outperforms regression-based directed evolution baselines on two in silico fitness landscapes.
-
Enhancing Safe and Controllable Protein Generation via Knowledge Preference Optimization
A knowledge-graph-guided preference optimization framework that fine-tunes protein language models to generate fewer sequences similar to known harmful proteins.
-
Evolution-Aware MSA Reasoning for Subsampling via Factor Graphs
AP-REASONER, an affinity-propagation factor-graph sampler with alpha/beta knobs, improves MSA-based protein LM pretraining for contact prediction and conformational sampling, though gains are modest and partly baselin...
Discussion (0). Sign in to comment.