Urgent.News

What's breaking now, across thousands of outlets.

AI

CytoGate-Bench: an LLM benchmark for cross-panel cell gating in cytometry

In cytometry, the workhorse single-cell technology of clinical immunology, every study defines its own antibody panel and cell-type vocabulary, so a classifier trained on one cannot annotate the next. Immunologists instead annotate by manual gating, splitting one parent population at a time on a two-marker plot, down an expert-defined hierarchy. We introduce CytoGate-Bench, a benchmark that…

Cytometry, a vital single-cell technology in clinical immunology, demands unique antibody panels and cell-type definitions for each study. Consequently, classifiers trained on one dataset cannot effectively annotate subsequent ones. Immunologists traditionally address this challenge through manual gating, sequentially partitioning populations based on multiple markers until a hierarchical structure is achieved.

To overcome this bottleneck, researchers have introduced CytoGate-Bench, a groundbreaking benchmark that transforms this iterative, step-by-step process into a zero-shot, panel-agnostic task for large language models. The benchmark consists of 23,646 expert-annotated instances, derived from 11 public flow- and mass-cytometry cohorts encompassing eight distinct marker panels.

When evaluated across six open- and closed-weight backbones, the most promising approach employs a single rectangular gate per candidate cell, matching the performance of trained, panel-specific baselines. Notably, the CytoGate-Bench formulation demonstrates greater resilience to distribution shifts compared to its counterparts.

Moreover, the benchmark reveals that navigating the hierarchical gating process stepwise yields superior results when compared to simultaneously predicting all cell types. This insight stems from analyses of data distribution shapes and curated marker priors, which prove instrumental in the performance enhancements.

Interestingly, further investigation points to the potential benefits of integrating vision capabilities and self-verification loops. These additions are found to systematically refine the gate boundaries, leading to more precise and accurate cell annotations.

Written by urgent.news from bioRxiv's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at biorxiv.org →

More in AI

What Actually Breaks When You Put AI Agents In Front Of Real Customers

I have spent the last year building AI automation systems for small businesses. Chatbots, multi agent workflows, the kind of stuff that looks great in a demo and then meets an actual customer who types "idk just fix it" and breaks everything. Most articles about AI agents talk about architecture.

  • Real users often provide contradictory, half-formed inputs that break AI agents.
  • Explicitly instructing the model to say "I don't know" when uncertain is crucial.
  • Maintenance is the most significant challenge due to frequent API and model changes.

More from Wednesday 26 August →