{
  "id": 4011004,
  "title": "Alibaba just released Qwen3.8-Flash: “An early preview of the architecture in Qwen4”",
  "url": "https://urgent.news/2026/08/28/alibaba-just-released-qwen3-8-flash-an-early-preview-of-the",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-28T18:27:06.000Z",
  "source": {
    "name": "The New Stack",
    "slug": "the-new-stack",
    "url": "https://thenewstack.io/qwen38-flash-previews-qwen4/"
  },
  "original_language": "en",
  "account": "Alibaba has recently launched Qwen3.8-Flash, an open-weight multimodal Mixture-of-Experts model, as an early preview of the Qwen4 architecture. This 125-billion-parameter AI model serves as a performance and value-for-money alternative to Qwen 4. By releasing Qwen3.8-Flash early, Alibaba aims to enable the community to examine and test the mechanics, constructs, and components of the architecture before the full Qwen4 model is developed.\n\nThe model is positioned to offer superior capabilities in coding and office tasks, while balancing capability, latency, and cost. Alibaba claims that Qwen3.8-Flash has a systematic improvement across four aspects: attention, residual, embedding, and optimization. This optimization leads to enhanced model capability, computational efficiency, model capacity, and training stability.\n\nBenchmarking Qwen3.8-Flash against rival models, such as DeepSeek-V4-Flash and Claude-Opus-4.6, shows that it performs well in agentic coding, long-horizon agent tasks, and multimodal intelligence. Qwen3.8-Flash-Next, the open-weight research frontier model, features a 125B parameter main model, supplemented by an additional 51B N-gram embeddings, and supports up to 1,000,000 tokens of context. Alibaba asserts that Qwen3.8-Flash requires only around one-ninth of the training resources of Qwen3.7-Plus, while delivering superior performance and significantly reducing both training and inference costs.",
  "summary": "Alibaba this week unveiled Qwen3.8-Flash, an open-weight, multimodal Mixture-of-Experts (MoE) model. Hot on the heels of Qwen 3.8 Max, which The post Alibaba just released Qwen3.8-Flash: “An early preview of the architecture in Qwen4” appeared first on The New Stack .",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}