GetChain News
中简 中繁 EN
GetChain News
Toggle sidebar

GLM-5.3-Flash Tops B.AI Model API Call Volume Rankings, Cumulative Throughput Surpasses 2.41 Trillion Tokens

Source: x.com Event types: Online/Update
The GLM-5.3-Flash model has become the most frequently invoked and popular model on the B.AI platform, with cumulative token throughput exceeding 2.41 trillion. As the first native multimodal model in the GLM-5 series, GLM-5.3-Flash features 320B total parameters and 18B active parameters. It employs a hybrid architecture combining sparse and linear attention mechanisms, supports 1M ultra-long context windows, and balances rapid response, powerful reasoning capabilities, and high cost-effectiveness. Starting today, developers can still invoke this model for free via the B.AI platform, covering diverse scenarios such as high-frequency APIs, coding, complex Agents, and ultra-long document processing. Try it now: chat.b.ai/chat

Related projects