ProtEnrich: Residual Multimodal Enrichment of Protein Sequence Embeddings
Protein language models effectively capture evolutionary and functional signals from sequence data but lack explicit representation of the biophysical properties that govern protein structure and dynamics. Existing multimodal approaches attempt to integrate such physical information through direct fusion, often requiring multimodal inputs at inference time and distorting the geometry of the…
We haven't written up this one. bioRxiv has the full story — the link below goes straight to it.