Gemini 3.7 Flash lands with coding gains at half the price
Google's Flash line is now on a three-week release cycle, and its newest iteration arrives with meaningfully better coding scores and a price tag that undercuts its own predecessor by 50 percent.
Google shipped Gemini 3.7 Flash on Thursday, three weeks after Gemini 3.6 Flash, calling it its "most intelligent workhorse model yet for coding and agents" — and pricing it at half of what the previous Flash charged at launch. The model scores 43.6 percent on Google's FrontierCode benchmark, up from 34.4 percent for 3.6 Flash, and 65.3 percent on the DeepSWE software-engineering benchmark versus 49.0 percent — results Google says put it ahead of both Claude Sonnet 5 and GPT-5.6 Terra on its own measurements. The gains extend beyond code: Google reports improvements in web development, document comprehension, and business process automation, and credits "awesome algorithmic improvements" rather than a bigger model for the jump in intelligence. The company also says the new Flash ships with stronger safeguards against misuse in the chemical, biological, radiological, and nuclear domains and in cyber offense.
The headline number, though, is the price. Gemini 3.7 Flash launches at $0.75 per million input tokens and $3.75 per million output tokens — exactly half of what 3.6 Flash cost when it debuted three weeks ago — and Google says the rate holds through the end of the year, with both Flash models now sharing the same price point. For agent builders, that math matters: agents burn tokens at a furious rate, and halving the marginal cost of a task is the difference between an agentic feature that is economically viable at scale and one that is not. The cuts also land on the same day DeepSeek pushed its own API prices up — we covered DeepSeek hikes API prices up to 4.7x with new peak-hour billing earlier today — a reminder that pricing pressure in this market runs in both directions.
The cadence is the other story. Three weeks between Flash generations points to a production line Google has clearly industrialized, and it is doing so while its flagship Gemini 3.5 Pro still hasn't shipped — the workhorse, not the flagship, is where Google is choosing to fight right now. Flash is the model most agentic apps will actually run on, and the company is betting that volume plus algorithmic efficiency beats the premium-tier arms race. The model is available through the API, AI Studio, and Antigravity, and consumer access arrives via Spark, Google's agent, for AI Ultra and Pro subscribers.
What to watch: whether OpenAI and Anthropic answer with price cuts of their own, and whether Google can hold a three-week Flash cadence through the rest of the year.
At $0.75 per million input tokens, does Flash's price-to-performance edge change which model you build on? Tell us in the comments.
Sources: Google DeepMind · The Decoder · Thurrott