AIToday
Large Language ModelsAI Coding AssistantsAI Business & IndustryTHE DECODERPublished: Aug 14, 2026, 04:01 JST3 min read

Google releases Gemini 3.7 Flash, cuts pricing 50% with coding gains

Google releases Gemini 3.7 Flash, cuts pricing 50% with coding gains

Key takeaway

  • Google shipped Gemini 3.7 Flash three weeks after its predecessor, delivering substantial gains on coding benchmarks—43.6% on FrontierCode versus 3.6 Flash's 34.4%—while undercutting the launch price by 50%.

  • The new model is now available through API, AI Studio, and Antigravity, priced at $0.75 per million input tokens and $3.75 per million output tokens.

3 Key Points

  1. What happened

    Google released Gemini 3.7 Flash, its successor to Gemini 3.6 Flash (released three weeks earlier), available through API, AI Studio, and Antigravity. The new model scores 43.6% on FrontierCode (up from 3.6 Flash's 34.4%) and 65.3% on DeepSWE (up from 49.0%), and Google says it outperforms both Claude Sonnet 5 and GPT-5.6 Terra on those benchmarks.

  2. Why it matters

    Gemini 3.7 Flash launches at $0.75 per million input tokens and $3.75 per million output tokens—50% cheaper than 3.6 Flash's launch price—while delivering notably stronger results on coding tasks, web development, document comprehension, and business process automation. For developers and businesses relying on AI for code generation and automation, this price-to-performance shift may reshape cost-benefit calculations.

  3. What to watch

    Google's pricing for both 3.6 Flash and 3.7 Flash holds through the end of the year; the company has indicated 'the models likely won't,' suggesting further changes are expected after that deadline.

In Depth

Read the full story

Google released Gemini 3.7 Flash on the heels of Gemini 3.6 Flash, which launched just three weeks prior. The company markets 3.7 Flash as "its most capable workhorse model yet for coding and AI agents," crediting what it calls "awesome algorithmic improvements" for the rapid leap in what Google describes as "intelligence."

On coding benchmarks, the gains are substantial. On FrontierCode, Gemini 3.7 Flash achieves 43.6%, a marked improvement over 3.6 Flash's 34.4%. On DeepSWE, the new model posts 65.3% versus 3.6 Flash's 49.0%. Google's own benchmark measurements also show improvements in web development, document comprehension, and business process automation. According to those same measurements, Gemini 3.7 Flash outperforms both Claude Sonnet 5 and GPT-5.6 Terra.

The model is immediately available through three channels: the API, AI Studio, and Antigravity. Pricing is set at $0.75 per million input tokens and $3.75 per million output tokens—50% cheaper than what 3.6 Flash cost at launch. Both models now carry the same price. Google notes that this pricing remains in effect through the end of the year, adding the remark that "the models likely won't," signaling that further product movement is expected once the year closes.

Context & Analysis

Google's three-week cycle from Gemini 3.6 Flash to Gemini 3.7 Flash underscores the company's aggressive cadence in releasing incremental improvements to its workhorse models. The company attributes the gains to "awesome algorithmic improvements," a notably informal framing that contrasts with typical enterprise software messaging—suggesting confidence in the underlying engineering rather than marketing narrative. The coding benchmarks tell a concrete story: on FrontierCode, a 9.2-percentage-point jump; on DeepSWE, a 16.3-percentage-point jump. These are substantial single-generation improvements, and the fact that Google's own measurements place the model ahead of both Claude Sonnet 5 and GPT-5.6 Terra lends credibility to the release as a genuine capability step forward, not a purely incremental refresh.

The pricing move is equally significant. By pricing 3.7 Flash at half the launch cost of 3.6 Flash—while improving performance—Google is signaling a shift in its economics for commodity inference tasks. The fact that both models "now share the same price point" is a direct competitive play, one that may reshape ROI calculations for developers and enterprises currently using or considering 3.6 Flash. The caveat that the pricing "holds through the end of the year; the models likely won't" is equally telling: it suggests both models are transitional products in a rapidly moving release cycle, and that pricing will reset once the next generation arrives.

FAQ

How much cheaper is Gemini 3.7 Flash than 3.6 Flash?
Gemini 3.7 Flash is 50% cheaper at launch, priced at $0.75 per million input tokens and $3.75 per million output tokens, compared to 3.6 Flash's higher launch price. Both models now share the same price point.
Where can I access Gemini 3.7 Flash?
The model is available through the API, AI Studio, and Antigravity.
How much better is Gemini 3.7 Flash at coding than its predecessor?
On FrontierCode, Gemini 3.7 Flash scores 43.6%, up from 3.6 Flash's 34.4%. On DeepSWE, it scores 65.3%, up from 3.6 Flash's 49.0%.

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleMicrosoft merges consumer and enterprise Copilot apps into super app

The AI news that matters, in one minute each morning.

Sign up free