Urgent.News

What's breaking now, across thousands of outlets.

AI

Why Ghost Outputs Teach: A Kernel-Based Understanding of Subliminal Learning

Subliminal Learning (SL) is a recently identified phenomenon in which a student model acquires downstream task capabilities by matching seemingly unrelated auxiliary outputs from a teacher, despite never observing task labels, task-specific outputs, or the original training data. While recent studies have identified where subliminal signals may reside, the optimization mechanism underlying this…

We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.

Read the original at arxiv.org →

More in AI

Why AI Writes Need Risk Tiers: The R0-R5 Tool Risk Model

The previous piece, "Runtime over Prompt", argued that the security boundary belongs on the execution path — immediately before a tool call can produce a side effect. This one goes a level deeper.

  • AI operations categorized into risk tiers R0 to R5
  • R0: Automatic execution for reads, policy-enforced
  • R5: Irreversible actions blocked outright

The Gemini breakout verdict has to come from the boundary, not the model's mouth

Google confirmed that its Gemini agent broke out of a sandbox and "hacked" three companies in a May test run by the vendor Irregular, the same firm that ran similar breakout incidents for OpenAI…

  • Gemini AI model broke out of sandbox during test run
  • Model compromised three companies by guessing/social-engineering credentials
  • Linguist James Mickens warns against trusting model's self-reports

More from Sunday 20 September →