[UTC+3]

Google Halves the Price of Gemini 3.7 Flash: Task Volume Gets Recalculated

August 21, 2026 · 1 minNewsReleases

On 13 August, Google introduced Gemini 3.7 Flash. It is an update to the workhorse class of models used for code and AI agents. The previous version shipped three weeks earlier.

Where has the cheap class improved?

Through the end of 2026, the price is $0.75 per 1M input tokens. Output runs $3.75 per 1M. That is half the launch price of Gemini 3.6 Flash.

Google's own benchmark numbers: FrontierCode 1.1 Main — 43.6% against 34.4%, DeepSWE v1.1 — 65.3% against 49%. The Elo rating in WebDev Arena — 1588 against 1538. On business processes in AutomationBench — 30.4% against 17%.

Gemini 3.6 Flash
Gemini 3.7 Flash
FrontierCode 1.1 Main
34.4%
43.6%
DeepSWE v1.1
49%
65.3%
WebDev Arena, Elo
1538
1588
AutomationBench, business processes
17%
30.4%
Google's own benchmark numbers · Source: itc.ua citing Google, 13 August 2026

Why redo the budget for high-volume work this quarter?

Document processing, code generation and support agents all run on the cheap class of models. Input costs half of what it did, so the same monthly volume comes out cheaper. The gain on AutomationBench points to fewer manual saves after an agent.

The math has to be run on real data. Take a month of actual traffic and run a sample of 200–300 tasks through both versions. Compare the share of tasks completed without human intervention. The inference bill and model switch calculators cover both sides of the equation.

What is still unknown?

The context limit, pricing after 31 December 2026, API and payment availability for Russian legal entities, and how long version 3.6 will be supported.

Source: itc.ua