Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection
Online log anomaly detection is critical for maintaining the reliability of large-scale computing systems. Although recent language model-based log anomaly detectors achieve strong detection performance, their confidence estimates remain poorly calibrated. We show that these detectors frequently assign excessive confidence to incorrect predictions, particularly for anomalous logs under severe…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.