Urgent.News

What's breaking now, across thousands of outlets.

AI

I built an AI incident responder that refuses to fix anything without asking

Built for the WeMakeDevs × TrueFoundry Agent Harness Hackathon. There are two kinds of "AI for incident response," and both of them are wrong. The first acts on its own. It sees latency spike, decides it knows why, and rolls back your deploy at 3am. When it's right, it's magic. When it's wrong — and it will be wrong, because production is where confident reasoning goes to die — you now have two…

Two types of AI for incident response were examined during the WeMakeDevs × TrueFoundry Agent Harness Hackathon. The first type independently takes action based on detected issues, which can lead to problems if it's incorrect. The second type simply notifies a human and provides log summaries, though this approach is both safe and ineffective. The goal was to create an AI incident responder that combines the speed of the first type with the safety of the second. The resulting solution is called Mayday.

Mayday is designed to prevent an AI from making changes without explicit human approval. When an alert is triggered, Mayday automatically scopes the incident, runs parallel analyses on metrics and logs, correlates with deployment history, and writes a diagnostic. It then proposes a single proposed fix, including root cause, expected impact, and what was ruled out.

The agent stops after proposing the fix, waiting for human approval before implementing any changes. A 401 HTTP error demonstrates that attempting to bypass this approval process is unsuccessful.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

Argentina most exposed to AI impact in Latin America, study finds

The risk is related to high urbanization and the prevalence of the service industry, with Buenos Aires leading the wa La entrada Argentina most exposed to AI impact in Latin America, study finds se…

  • Argentina ranks highest in Latin America for AI exposure at 0.292
  • Service sector's 88% employment drives Argentina's AI vulnerability
  • AI adoption increases demand for supervision and judgment skills

This is what every AI pin gets WRONG

If you are interested to get this project developed further, cosider starring it on my github: life-autopilot-agentic WITH the way things stand, AI pins promised the FUTURE, but the technology simply…

  • AI pin failed to meet consumer and investor expectations
  • Device's small battery caused overheating issues
  • Laser projector touted as impressive feature but lacked practical utility

The #1 row on this AI memory leaderboard is not a measurement

Bench'd (benchd.ai) calls itself the neutral benchmark authority for AI memory, and sells vendors a verification badge from $299 to $3,999.99 a month. I ran my memory system through their harness.

  • AI memory leaderboard's top entry is not genuine measurement
  • Benchmark service Bench d sells verification badge for $299-$3,999.99/month
  • Top three entries lack downloadable proof of Community-Verification

More from Saturday 29 August →