Google Gemini 2.5 Flash Reaches General Availability Across Search, APIs, and Enterprise
Google officially launched Gemini 2.5 Flash as a generally available model on August 13, 2026, deploying it across the Gemini API, Google AI Studio, Vertex AI, Gemini Enterprise, the Gemini app, and AI Mode in Search. The model succeeds earlier Flash generations and emphasizes stronger instruction following, better intent understanding, and faster responses for coding, agentic workflows, and multi-step tasks. It supports a 1 million-token context window, up to 64,000 output tokens, and adjustable thinking levels, making it suited for workloads requiring large-scale text processing or variable reasoning depth. In AI Mode in Search, Gemini 2.5 Flash is replacing earlier Flash variants for many users in supported markets, while Google AI Pro and Ultra subscribers can access it via Gemini Spark as the rollout continues. Google has set introductory pricing at $0.75 per million input tokens and $3.75 per million output tokens through December 31, 2026, after which standard pricing will apply.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in