Urgent.News

What's breaking now, across thousands of outlets.

AI

Principled Under Pressure: Post-Training Decides Whether LLMs Act on Their Own Moral Judgment

Language models increasingly act as agents. An agent that says an action is wrong and then takes it anyway is a different failure from one that does not know better, and evaluations of stated values cannot see it. We build a pre-registered panel of 248 scenarios across five kinds of pressure. Each scenario is posed twice to the same model, once as the agent choosing what to do and once in the…

We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.

Read the original at arxiv.org →

More in AI

Robin: the AI that designs experiments but does not carry them out

Un robot no eligió solo la molécula correcta para tratar una enfermedad incurable del ojo: la propuso, y después un laboratorio real confirmó que funcionaba.

  • Robin, a robot orquestador, proposed drug ripasudil for age-related macular degeneration in 2025
  • Predecessors Adam and Eve worked on genetics of yeast and antimalarials in 2009 and 2015

More from Tuesday 6 October →