Z.ai Reveals GLM-5.3-Flash: 100,000-Chip Domestic Cluster Achieves 3x Performance Gain
On September 17, 2026, Z.ai published a technical blog revealing that GLM-5.3-Flash, previously tested anonymously as Ox Alpha, now runs entirely on more than 100,000 domestic AI accelerators. It achieves triple the performance of the initial baseline with per-token costs comparable to mainstream solutions.