Zhipu AI shares jump as viral Ox Alpha model revealed as GLM-5.3-Flash on Chinese chips
China’s Zhipu AI has launched its latest open-weight model, GLM-5.3-Flash – previously code-named Ox Alpha – saying that the system ran entirely on a cluster of 100,000 domestically produced chips during a high-profile stealth trial. The announcement followed a week of heavy traffic on artificial intelligence model marketplace OpenRouter and agent platform OpenCode, where the model processed 62…
China's Zhipu AI introduced its newest open-weight model, GLM-5.3-Flash, also known as Ox Alpha, during a secret trial utilizing 100,000 locally manufactured chips. The model processed 62 trillion tokens during its trial, which was the largest launch on the OpenRouter marketplace. Zhipu's shares rose over 12% in Hong Kong after the announcement.
Despite heavy traffic during the free trial, the model's output speed was slower than the industry average. The GLM-5.3-Flash boasts 320 billion total parameters and can process visual information alongside text. It is the first in its series to perform this task. While Zhipu has not disclosed the chip suppliers for the cluster, the company has previously worked with international developers like Huawei, Cambricon, and Moore Threads.
The model offers significant cost savings for international developers, with prices around 1/10th of standard GLM-5.3. Zhipu has made the model weights globally available and integrated GLM-5.3-Flash into its ZCode platform and GLM Coding Plan. This launch comes amid increasing competition in China's open-source AI sector, with Alibaba releasing Qwen3.8-Flash-Next, a multimodal preview of Qwen4.
Written by urgent.news from SCMP Tech's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.
This story
This is one outlet's version. Read the fullest account.
- Zhipu AI shares jump as viral Ox Alpha model revealed as GLM-5.3-Flash on Chinese chips scmp.com
- Z.ai open-sources ‘Ox Alpha’ model as GLM-5.3-Flash siliconangle.com
- Alibaba releases Qwen3.8-Flash, an open-weight, 125B-parameter model built on its next-gen Qwen 4 architecture, saying it rivals Opus 4.6 and V4-Flash (Luz Ding/Bloomberg) bloomberg.com
- Z.ai releases GLM-5.3-Flash, the first natively multimodal GLM-5 series model, with 320B parameters, saying it outperforms GLM-5.2 at "one-tenth the price" (Z.ai) z.ai
- Alibaba's Qwen launches Qwen3.8-Flash AI model with lower training costs economictimes.indiatimes.com