Gemini Dual Models: 3.6 Flash and 3.5 Flash-Lite Now Available on B.AI API
Odaily Planet Daily News On July 22, B.AI API officially launched two models: Gemini 3.6 Flash and Gemini 3.5 Flash-Lite. Among them, Gemini 3.6 Flash, as an upgraded version of 3.5 Flash, significantly improves output quality while maintaining the same pricing, reducing token consumption by an average of approximately 17%. Notably, in complex DeepSWE scenarios, the reduction can reach up to 65%, greatly enhancing development efficiency and cost-effectiveness. Gemini 3.5 Flash-Lite, with its blazing fast output speed of up to 350 tokens/s and extreme cost-performance ratio, precisely meets the demands of high-concurrency, high-real-time tasks such as document batch processing and Agentic intelligent retrieval. Both models now have their official API interfaces open, allowing developers to log in to the B.AI platform immediately for seamless access and experience, fully leveraging the performance dividends brought by the new generation of models.
