Gemini Dual Models 3.6 Flash and 3.5 Flash-Lite Land on B.AI API
B.AI API has officially launched two models: Gemini 3.6 Flash and Gemini 3.5 Flash-Lite. Among them, Gemini 3.6 Flash, as an upgraded version of 3.5 Flash, significantly improves output quality while maintaining the same pricing, with Token consumption reduced by approximately 17% on average; especially in complex DeepSWE scenarios, the reduction can reach up to 65%, greatly enhancing development efficiency and cost-effectiveness. Gemini 3.5 Flash-Lite delivers ultra-fast output of up to 350 tokens/s and ultimate cost-effectiveness, precisely meeting the demands of high-concurrency, high-real-time tasks such as document batch processing and Agentic intelligent retrieval. Currently, official API interfaces for both models are available. Developers can log in to the B.AI platform immediately to seamlessly access and experience them, enjoying the performance benefits brought by the new generation of models.