Urgent.News

650+ sources. One page. See who else covered it.

Editions

AI

I Asked the Same Question to 7 Local LLMs — Speed and Intelligence Didn't Line Up: DGX Spark Benchmarks

Originally published on my Substack . I'm a Microsoft MVP based in Japan, writing in English about the AI agent systems I actually run in production. Local AI models keep multiplying. But comparing numbers on model cards alone doesn't tell you which one to actually use. Does a higher parameter count mean smarter? Does MoE mean faster? If a model is popular on AI Arena, is it good for my own work?…

I conducted a test using the same query across seven major language models (LLMs) running on a single NVIDIA DGX Spark system. The goal was to evaluate not just the speed but also the quality and reliability of the generated answers. Among the models tested, Qwen3.5 35B was the quickest, completing the task in 1.49 seconds. However, its answer contained a risky claim about zero risk of confidential data leakage, which is not suitable for business use.

On the other hand, Qwen3.6 35B-A3B, which took slightly longer at 2.12 seconds, provided a more practical answer that addressed the risk of data leakage and required specialized knowledge for setup. Ultimately, while speed is an important factor, it did not consistently correlate with the reliability or suitability of the answers for business applications.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

I Built a World Where the Canon Is Written by AI Agents — 13 Artifacts, 5 LLMs, 0 Human Gatekeepers

Cover story Not a prompt library. Not a chatbot wrapper. A 1,000-year future history (2025–3000+) where the world itself is a set of hard rules, and the content is written by whichever LLM decides to…

  • Thirteen artifacts created by five LLMs, no human gatekeepers
  • Four autonomous writing schools emerge without human design
  • Contradictions between documents accepted as part of canon

Claude Impact Lab LA: Community Changed the Code

Eighty minutes into building with three people I had met that morning, I renamed the idea I brought with me. 21:05 Rename the product to Civiq and credit the team I wrote that commit message myself.

  • Claude Impact Lab LA transformed community ideas into functional product in one day
  • Agenda Watch tool created to make city government agendas searchable and transparent
  • Participants collaborated across skill levels, one overcame shyness through group setting

Stop Sending Your Vitals to the Cloud: Running Llama-3 Locally in the Browser with WebLLM & WebGPU 🥑

Privacy is the ultimate "final boss" in HealthTech. When users record sensitive medical logs, the last thing they want is their data being used to train a massive corporate model.

  • Developers run Llama-3 model locally in web browsers using WebGPU.
  • Privacy-first health apps process user inputs without transmitting PHI.
  • System stores structured health data locally in browser's IndexedDB.

More from Sunday 16 August →