Zhipu AI has introduced the GLM-5.3-Flash model, which, according to the company, operates exclusively on domestic artificial intelligence chips. One hundred thousand locally produced accelerators are used to process all of the model's online traffic.
The model ranked tenth in the AAII rating from the analytical firm Artificial Analysis, surpassing DeepSeek's V4 Pro Max. Zhipu AI emphasized that the entire volume of requests to GLM-5.3-Flash is served by these 100,000 domestically manufactured chips.
GLM-5.3-Flash was initially released under the codename 'Ox Alpha' on August 20th and quickly topped the popularity lists on the OpenRouter AI model routing platform within a week. The company positions this model as high-performance, sufficiently efficient for large-scale operation on China's internal computing infrastructure.
Zhipu AI has not publicly disclosed the specific chip supplier used. However, according to CNBC reports, analysts suggest that the hardware belongs to Huawei's Ascend series. Ivan Lin, an analyst at the research firm Counterpoint, noted that developers of Chinese AI models are increasingly directing investments into AI servers and computing infrastructure built on domestic chips.
This launch is the clearest signal yet that at least one leading model development laboratory in China believes that the domestic supply chain is ready to support flagship workloads. Zhipu AI is among the most active 'Chinese tigers' in expanding computing power, and the daily operation of a frontier-class model on its own accelerators is a statement about both performance and chip availability.
This move also aligns with a broader industry trend. As export controls restrict access to advanced foreign chips, creators of Chinese models are rushing to optimize their software for the hardware they can actually acquire. A model demonstrating high results in benchmarks while running on domestic chips serves as proof that software efficiency can compensate for hardware gaps.
For Zhipu AI, the Flash line represents a strategy combining model quality with practical scalability. With the support of 100,000 chips, GLM-5.3-Flash aims to demonstrate that the Chinese laboratory is capable of providing competitive frontier-level performance without reliance on imported accelerators.



