POLAR: Ontology-Guided Risk Prevention for Tool-Calling LLM Agents
LLM tool-use agents operate in dynamic environments where many actions carry operational risk. However, most safety mechanisms react only after errors manifest. Existing pre-emptive approaches either fine-tune the agent on chain-of-thought deliberation or compile natural-language guardrails into runtime checks, but they do so without exposing a structural, auditable verdict. We propose POLAR, a…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.