Automating scientific annotations for open transcriptomic profiles via multi-stage agents
Public transcriptomic repositories contain millions of samples, yet their large-scale reuse is hindered by heterogeneous and inconsistently reported metadata. In the Gene Expression Omnibus (GEO), key biological information is often distributed across study- and sample-level records, requiring context-dependent interpretation. Here we present GEOMeta, a large language model (LLM)-based…
We haven't written up this one. bioRxiv has the full story — the link below goes straight to it.