{
  "id": 10734440,
  "title": "DisenTE: Sparse Pattern-Context Modeling for Interpretable Translation-Efficiency Matrix Completion",
  "url": "https://urgent.news/2026/09/29/disente-sparse-pattern-context-modeling-for-interpretable-translation",
  "topic": "science",
  "section": "Science",
  "published": "2026-09-29T00:00:00.000Z",
  "source": {
    "name": "bioRxiv",
    "slug": "biorxiv",
    "url": "https://www.biorxiv.org/content/10.64898/2026.09.23.753721v1?rss=1"
  },
  "original_language": "en",
  "account": "In the realm of data-rich science, object-by-context matrices often present partially observed data, where dominant object effects can mask smaller but significant context-dependent variations. This challenge was examined in a translation-efficiency atlas of 9,494 5 UTRs spanning 78 cellular and tissue contexts. The researchers introduced DisenTE, a neural model capable of handling these complexities.\n\nDisenTE employs a sequence-conditioned architecture, integrating distinct sequence and context branches linked by a sparse low-rank pattern-context channel. Each module merges a sequence-derived activation with context-specific deployment weights, yielding a dictionary where both sequence and context components can be analyzed independently.\n\nWhen subjected to five-fold within-panel entry masking, DisenTE demonstrated an impressive UTR-centered residual Spearman correlation of 0.641 +/- 0.005, a marked improvement over the 0.304 +/- 0.003 achieved by the best reference model. The learned dictionary preserved 11 out of 20 candidate modules, including CTM 6, which exhibited the strongest overlap with an external TOP set and a cap-proximal pyrimidine pattern. CTMs 5 and 7 also shared similarities with the set but featured purine-containing consensuses, suggesting their potential as TOP-set-associated factors.\n\nThe findings bolster CTM 6's status as a top sequence anchor and position CTMs 5 and 7 as key elements of the TOP set. Importantly, DisenTE outperformed the evaluated references, completing the dataset with greater accuracy while concurrently generating module-level summaries of its context-dependent variations.",
  "summary": "Partially observed object-by-context matrices arise across data-rich science, where dominant object effects can obscure smaller but informative context-dependent variation. We study this problem in a translation-efficiency atlas of 9,494 5' UTRs across 78 cellular and tissue contexts. We present DisenTE, a sequence-conditioned neural model that combines separate sequence and context branches with…",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}