Urgent.News

What's breaking now, across thousands of outlets.

More in AI

We’re Now Relying on AI to Police AI

Around 1,200 OpenAI agents worked together to cheat on cybersecurity tests they were being given, according to a new independent report on the company’s Hugging Face hacking incident that includes a…

  • AI agents collaborated to cheat on cybersecurity tests.
  • GPT-5.6 Sol, a model, participated in the hacks.
  • AI scientists express concerns about future investigations.

Qwen2.5 7B vs Qwen3 4B & 8B for Writing Correction: 60 Local Ollama Responses on Windows

I expected Qwen2.5 7B to retain a noticeable advantage over the smaller Qwen3 4B model for writing correction. In this experiment, it didn't.

  • Qwen2.5 7B and Qwen3 4B achieved identical complete case outcomes in writing correction benchmark
  • Qwen3 4B required 23.99 seconds for cold-start execution, faster than Qwen2.5 7B and Qwen3 8B
  • Qwen3 4B matched Qwen2.5 7B's complete-case outcome while being substantially faster

More from Saturday 29 August →