Urgent.News

600+ sources. One page. See who else covered it.

Editions

AI

Agentic reliability and evaluations : Enterprises that got burned by a bad eval are the most likely to remove humans from the loop, not the least

Across 108 enterprises, trust in automated agent evaluation rose sharply in July — and the failure rate it is supposed to predict did not move at all. The share of organizations that fully trust automated evaluation nearly tripled, from 5% in June to 13%, and the complaint that evaluations don’t match real-world outcomes fell 10 points. Yet the same share as last month — just under half — shipped…

Agentic reliability and evaluations : Enterprises that got burned by a bad eval are the most likely to remove humans from the loop, not the least

We haven't written up this one. VentureBeat has the full story — the link below goes straight to it.

Read the original at venturebeat.com →

More in AI

Google's Gemini AI assistant reaches one billion users

Google's Gemini AI assistant reaches one billion users

SAN FRANCISCO: Google’s artificial intelligence assistant Gemini has surpassed one billion monthly users, CEO Sundar Pichai announced Tuesday, a week after a major reorganization of the tech giant’s AI division. “It’s our fastest growing product ever,” and the 14th product to reach a billion users, Pichai posted on social media.

Telecoms see AI businesses gain traction in Q2

Telecoms see AI businesses gain traction in Q2

Major telecom operators are seeing early results from their push into artificial intelligence (AI), with AI data centers (AIDC), cloud and AI transformation businesses posting growth in the second quarter even as traditional telecom operations remained largely stagnant.

More from Wednesday 12 August →