Google Registers Gemini 3.6 Flash and 3.5 Flash Lite AI Models

iconCryptoBriefing
Share
AI summary iconSummary
AI + crypto news: Google has registered Gemini 3.6 Flash and 3.5 Flash Lite AI models internally, with model IDs appearing in development tools. The models may soon launch via AI Studio and Vertex API. The 3.6 Flash model shows signs of tiered pricing. No official project announcement has been made as of July 21, 2026. The Gemini 3.5 Pro remains delayed, with Flash and Flash Lite acting as temporary options.

Google appears to be prepping two new AI models for deployment, with identifiers for Gemini 3.6 Flash and Gemini 3.5 Flash Lite surfacing inside the company’s development infrastructure. The model ID gemini-3.6-flash-tiered was spotted in Google’s Antigravity IDE, suggesting the models are at least close to being ready for external access through AI Studio and the Vertex API.

What we actually know about the new models

The Gemini 3.6 Flash and 3.5 Flash Lite registrations were first noted on July 21, 2026. Internal registration in development tools is typically one of the final steps before a model becomes publicly accessible, though it doesn’t guarantee an imminent launch.

Google’s existing Gemini 3.5 Flash model remains operational on both AI Studio and Vertex AI. That model currently runs at approximately $1.50 per million input tokens and $9 per million output tokens, with a context window of 1 million tokens.

Advertisement

The “tiered” suffix in the registered ID for the 3.6 Flash model hints at a pricing structure with multiple access levels, possibly differentiating between rate-limited free tiers and paid production tiers.

Neither model has received a formal public announcement from Google, and no documentation has appeared on the company’s developer pages as of this writing.

The Gemini 3.5 Pro problem

Google’s Gemini 3.5 Pro has been conspicuously delayed. The Pro model represents Google’s top-tier reasoning and capability offering, designed to compete with the most advanced models from OpenAI, Anthropic, and other frontier labs. Releasing Flash and Flash Lite variants as interim solutions while the Pro model faces delays is consistent with Google’s historical pattern of pairing rapid Flash-tier releases with more refined Pro models.

What this means for the AI market and investors

Google parent Alphabet remains one of the largest companies investing in AI infrastructure and model development globally. The registration of new model variants signals continued commitment to expanding the Gemini ecosystem, even as the flagship release timeline slips.

At roughly $1.50 to $9 per million tokens for the current Flash model, Google is positioning itself competitively against rivals. A Flash Lite model could push those costs even lower, potentially undercutting competitors who charge premium rates for comparable inference quality.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.