B.AI Launches Gemini 3.6 Flash and 3.5 Flash-Lite; Token Consumption Down 17%

According to B.AI, Gemini 3.6 Flash and Gemini 3.5 Flash-Lite are now available on the B.AI API as of July 22. Gemini 3.6 Flash, an upgraded version of Gemini 3.5 Flash, reduces token consumption by approximately 17% on average while maintaining the same pricing, with reductions reaching up to 65% in complex scenarios like DeepSWE.

Gemini 3.5 Flash-Lite delivers throughput of 350 tokens per second and offers extreme cost efficiency for high-concurrency tasks including document batch processing and agentic retrieval. Both models are now available via official API endpoints.

Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments