News linked to both this project and an event.
The B.AI platform announced that Z.AI's speed-optimized native multimodal large model GLM-5.3-FlashX has officially launched, with both the API and Web Chat now open for public testing.
Zhipu has officially announced the launch of GLM-5.3-FlashX. The model was previously made available to global developers under the name "Ox Alpha". According to Zhipu, leveraging inference compute from 100,000 domestic chips, the company has further increased infrastructure investment and optimized inference performance, achieving speeds of up to 200 Tokens/s. The GLM-5.3-FlashX API is now live, with the model identifier set to "GLM-5.3-FlashX".