Google Launches Gemini 3.6 Flash, Delays Gemini 3.5 Pro Release

icon币界网
Share
AI summary iconSummary
Google has launched Gemini 3.6 Flash, reducing token usage by 17% and pricing output at $7.5 per million tokens. The model is now available via the Gemini app, Google AI Studio, and API. New token listings are anticipated as the update supports enhanced on-chain news integration. Meanwhile, Gemini 3.5 Pro remains in internal testing due to performance issues, particularly in programming tasks. Gemini 4 has begun pre-training, signaling continued development.
CoinDesk reports:

Google has launched three new Gemini models, with key updates focused on invocation costs, response speed, and agent-based AI scenarios. However, the much-anticipated Gemini 3.5 Pro has not yet been released, indicating a delay in the product roadmap.

3.5 Pro remains in the testing phase.

At Google's I/O 2026 event in May, the company stated that Gemini 3.5 Pro would be released within a month. However, according to reports, the model failed to meet its targets during internal testing, particularly underperforming on programming tasks, and thus was not launched on schedule.

Bloomberg previously reported that Google attempted to address the issue at the end of June by updating its training data, but the results were unsatisfactory. Google has so far only stated that Gemini 3.5 Pro will be released “when ready.”

3.6 Lower Prices and Improve Efficiency on Flash

The core release of this update is Gemini 3.6 Flash, which continues to be positioned as a lightweight model designed for AI agents and high-frequency use cases. Compared to Gemini 3.5 Flash, this model reduces output token usage by approximately 17% and lowers the output price from $9 to $7.5 per million tokens.

Public benchmark results show that Gemini 3.6 Flash demonstrates significant improvements over its predecessor across multiple tasks, including software engineering, machine learning engineering, and computer operation tasks. Specifically, the DeepSWE v1.1 score increased from 37% to 49%, MLE-Bench rose from 49.7% to 63.9%, and OSWorld-Verified reached 83.0%.

However, in certain programming and knowledge-based tests, OpenAI and Anthropic’s comparable models still lead. The report also noted that Gemini 3.6 Flash still encountered formatting errors and produced results that were not directly usable in simple programming benchmarks.

Flash-Lite and Cyber are designed for different use cases.

Gemini 3.5 Flash-Lite emphasizes high throughput and low cost, making it ideal for document processing, search systems, and batch workflows. This model can output up to 350 tokens per second, with an input price of $0.30 per million tokens and an output price of $2.50 per million tokens.

Another version of Gemini 3.5 Flash Cyber will not be available to the public. Google has restricted access to government agencies and vetted partners for the purpose of identifying and fixing software vulnerabilities.

Gemini 4 has started pre-training.

Google has also confirmed that pre-training for Gemini 4 has begun. Pre-training is the stage in which a model learns foundational data at scale, indicating that Gemini 4 has entered the actual development process, rather than remaining in the planning phase.

Gemini 3.6 Flash and Gemini 3.5 Flash-Lite are now available on the Gemini app, Google AI Studio, and API.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.