❯ Google Launches Gemini 3.7 Flash, Cuts Intro Price in Half, Targets Coding and Agents
LEADGoogle launched Gemini 3.7 Flash on August 13, just 3 weeks after 3.6 Flash debuted. Intro pricing is $0.75 per million input tokens and $3.75 per million output tokens — half the prior generation’s launch price. The rate holds through the end of the year, then reverts to $1.50 and $7.50 in 2027. Product lead Tulsee Doshi called it Google’s “most intelligent flagship model,” aimed at coding and agent use cases.
IMPACTThe gains are concentrated in coding and automation. Per official results: DeepSWE v1.1 rose from the prior generation’s 49.0% to 65.3%, FrontierCode 1.1 improved from 34.4% to 43.6%, and AutomationBench nearly doubled, from 17.0% to 30.4%. Third-party benchmark firm Artificial Analysis gave it an intelligence index of 56 — 4 points higher than the model released three weeks ago — and placed it on the “intelligence vs. latency” Pareto frontier.
CONTEXTOver the past 3 months, Google had already shipped two Flash models; this is the third. Developer Simon Willison flagged the awkward pricing design: the intro price is set to double on December 31, yet at a three-week iteration cadence, who will still be using this generation five months from now? The answer, in all likelihood, is no one. Google only wants today’s adoption.
DETAILWhat gets squeezed is the mid-tier price band: Claude Sonnet, GPT-5.6 Terra, and a host of open-source models in the same class all need to justify why they cost more. Teams building agent applications will have to recalculate their per-task cost comparison tables this week.
▪ SIGNALA release every three weeks, a price halved with each release — Google is wielding iteration speed as a weapon, turning mid-tier model margins into a war of attrition.