Google releases Gemini 3.7 Flash with discounted agent pricing through 2026
Google released Gemini 3.7 Flash as a generally available model for coding and agents, with introductory API pricing running through December 31, 2026.
AI-generated: This article was created and published automatically by LinkLoot and was not substantively reviewed by a human editor.
Google releases Gemini 3.7 Flash with discounted agent pricing through 2026
Google released Gemini 3.7 Flash on August 13, 2026, positioning it as its most capable Flash workhorse model for coding, web development, knowledge work, and agent workflows. The model is generally available in the Gemini API as gemini-3.7-flash.
The pricing detail is time-sensitive: Google says the introductory price is $0.75 per million input tokens and $3.75 per million output tokens through December 31, 2026. After that window, the standard listed price rises to $1.50 per million input tokens and $7.50 per million output tokens.
Key takeaways
- Gemini 3.7 Flash is now generally available in the Gemini API under the stable ID
gemini-3.7-flash. - Google says it improves coding, web development, knowledge work, and agentic workflows over Gemini 3.6 Flash.
- The model supports text, image, video, audio, and PDF inputs, with text output.
- Introductory API pricing runs through December 31, 2026, according to Google’s model documentation.
What launched
Google describes Gemini 3.7 Flash as the next iteration in the Gemini 3 Flash series and says it was shaped by developer feedback and algorithmic changes. The official blog emphasizes software engineering, web development, enterprise workflow automation, and document-heavy knowledge work rather than consumer chat features alone.
The Gemini API changelog lists the model as generally available, which makes this more than a teaser or early-watch item. Developers can build against the stable model ID and consult the current Gemini model page for limits and supported data types.
Where the model fits
Gemini 3.7 Flash keeps the Flash positioning: a workhorse model meant to balance quality, speed, and cost for high-volume production use. Google’s model page lists a 1,048,576-token input limit and a 65,536-token output limit, with text, image, video, audio, and PDF inputs.
The model is also tied to Gemini Spark, Google’s personal AI agent surface for Pro and Ultra subscribers in more than 160 countries. Google says Spark will use 3.7 Flash to improve complex workflows across files, email, and status documents.
Pricing and the deadline
The pricing window is one of the most actionable parts of the release. Google’s blog and model documentation list discounted pricing through the end of 2026. Artificial Analysis also highlights the discount and compares cost per task against Gemini 3.6 Flash.
For teams evaluating agent backends, that makes the next few months a useful test period. The cheaper price may justify heavier benchmark runs, but production budgets should account for the January 1, 2027 standard pricing if Google does not extend the discount.
Independent benchmark context
Artificial Analysis says Gemini 3.7 Flash reaches its Intelligence vs. Time per Task Pareto frontier and reports a four-point improvement over Gemini 3.6 Flash on its Intelligence Index. It also reports strong performance on agentic and spreadsheet/document tasks.
Those external numbers should not replace internal testing. They do, however, support Google’s claim that the release is aimed at practical coding and agent work rather than a narrow benchmark-only update.
Source check
- Google announcement confirms the launch, target workflows, Gemini Spark rollout, and introductory pricing.
- Gemini API changelog confirms general availability and the stable
gemini-3.7-flashmodel ID. - Gemini 3.7 Flash model page lists supported inputs, token limits, and model status.
- Artificial Analysis provides independent benchmark and cost-per-task context.
Try the related loot
Debug Cloudflare Workers locally with traces an AI agent can read
