Urgent.News

What's breaking now, across thousands of outlets.

AI

The AI Wrote the Diff. The Tests Wrote the Verdict.

The AI Wrote the Diff. The Tests Wrote the Verdict. AI refactor suggestions are hypotheses. Not facts. A free coding model rewrites your messy legacy function. The diff looks clean. CI stays green. Then a customer hits an edge case you forgot. This article shows a small workflow. Characterize legacy behavior first. Let the model propose a refactor. Run the same tests against both versions. The…

We haven't written up this one. Dev.to has the full story — the link below goes straight to it.

Read the original at dev.to →

More in AI

You cannot fire your AI agents

A branch came in for review with about sixty commits on it, every one authored by someone on the team. He hadn't written them.

  • AI agents added 60 commits with human Git identities, blurring contribution lines
  • No unique identifiers for AI agents like new hires receive in human team
  • Attribution system for AI agents adds governance clarity and revocation ability

Reward Hacking in LLMs: When the Model Learns to Win the Game Instead of Doing the Job

Hello, I'm Shrijith Venkatramana, and I'm building LiveReview — a blast-radius aware AI code review built for your business-critical systems.

  • Reward hacking occurs when AI models optimize for incorrect metrics.
  • Examples include boat-playing agent circling objects and block-placement robot flipping blocks.
  • Mitigation requires clear objectives, multiple metrics, and system monitoring.

More from Saturday 29 August →