Zhipu GLM Team Reports Recursive Self-Improvement Deployment on 100,000-Plus Domestic Chip Cluster
Zhipu's GLM team disclosed its first engineering deployment of recursive self-improvement on September 17, 2026. An Infra Agent driven by GLM-5.3 built and debugged inference services from scratch on a cluster of more than 100,000 domestic chips, tripling overall processing speed in under two weeks. Hardware utilization and per-token cost approached Nvidia GPU levels. GLM-5.3-Flash, launched anonymously as Ox-Alpha, drew over 620 trillion token calls in six days.
On September 17, 2026, the Zhipu GLM team disclosed its first engineering progress on recursive self-improvement. An Infra Agent driven by GLM-5.3 built and debugged inference services from scratch on a cluster comprising more than 100,000 domestic chips, raising overall processing speed to three times the original level in under two weeks, with hardware utilization efficiency and per-token cost approaching Nvidia GPU levels.
The system has been put into real use. GLM-5.3-Flash was launched as the anonymous model Ox-Alpha, and its token call volume exceeded 620 trillion within six days. This is the first time a domestic large-model vendor has publicly applied AI self-improvement in a production environment.
Why this event matters
The event has a measured impact on 3 industrys. The strongest current signal is positive for Artificial Intelligence, with intensity 85/100 and 75% confidence over a medium term horizon.
Artificial Intelligence
- Direction
- positive
- Intensity
- 85
- Confidence
- 75%
- Horizon
- Medium term
Semiconductor Value Chain
- Direction
- positive
- Intensity
- 75
- Confidence
- 70%
- Horizon
- Medium term
Cloud Services & Data Centres
- Direction
- positive
- Intensity
- 65
- Confidence
- 65%
- Horizon
- Short term
Impact figures are analytical estimates that combine direction, intensity, confidence and event importance. They are not investment advice.