ProtJEPA: A Multimodal Joint-Embedding Predictive Architecture for Protein Biological World Modeling with Multi-TeacherModality-Attentive Fusion
Over 99.9% of known protein sequences lack experimentally validated functional annotations. We present ProtJEPA, a multimodal Joint-Embedding Predictive Architecture that trains a sequence-only student encoder to predict joint embeddings spanning ten biological modalities - sequence, structure, knowledge graph, protein interactions, literature, localization, tissue expression, GO function,…
We haven't written up this one. bioRxiv has the full story — the link below goes straight to it.