Zhipu officially launches GLM-5.3-FlashX
Zhipu has officially announced the launch of GLM-5.3-FlashX. The model was previously made available to global developers under the name "Ox Alpha". According to Zhipu, leveraging inference compute from 100,000 domestic chips, the company has further increased infrastructure investment and optimized inference performance, achieving speeds of up to 200 Tokens/s. The GLM-5.3-FlashX API is now live, with the model identifier set to "GLM-5.3-FlashX".