Skip to content
llm-spend
GitHub

Gemini

Google

Gemini 3.7 Flash is the new flagship, and Gemini 3.6 Flash's price is halved to match it — both $0.75/$0.075/$3.75 per M through year-end.

Gemini 3.7 Flash is now the catalog's flagship Gemini model for agentic and multimodal work. The same pricing update halved Gemini 3.6 Flash to match it exactly: both now $0.75/M input, $0.075/M cached, $3.75/M output, a promotional rate published through 2026-12-31 that reverts to $1.50/$0.15/$7.50 on 2027-01-01. Gemini 3.5 Flash remains listed for comparison, unchanged at $1.50/$0.15/$9.00.

Pricing

per 1M tokens · USD / CHF
ModelTier / HostInputCachedOutputConfidence
Gemini 3.7 Flash
New flagship Flash model; same promotional structure as 3.6 Flash — $0.75/$0.075/$3.75 through 2026-12-31, reverting to $1.50/$0.15/$7.50 on 2027-01-01. Priced identically to 3.6 Flash on every dimension.
Foundry ·Global$0.75CHF 0.604$0.075CHF 0.06$3.75CHF 3.02official
Gemini API pricing page (ai.google.dev/gemini-api/docs/pricing), captured 2026-08-14; page stamped "Last updated 2026-08-13 UTC", same page as the 3.6 Flash row below. Google describes it as "Our most capable Flash model for agentic workflows and multimodal reasoning." The reversion date is published inline per price cell, verbatim: "$0.75 through December 31, 2026.$1.50 starting January 1, 2027." Batch/Flex bill at exactly 50% of Standard and Priority at exactly 1.8x, in both the current and post-reversion periods — now modeled as service-tier variants below.
Standard (from 2027) begins 1 Jan 2027 UTC
Standard (from 2027)$1.50 / $0.15 / $7.50Batch$0.375 / $0.037 / $1.88Batch (from 2027)$0.75 / $0.075 / $3.75Flex$0.375 / $0.037 / $1.88Flex (from 2027)$0.75 / $0.075 / $3.75Priority$1.35 / $0.135 / $6.75Priority (from 2027)$2.70 / $0.27 / $13.50
Gemini 3.6 Flash
Promotional rate (50% off standard pricing) through 2026-12-31; reverts to $1.50/$0.15/$7.50 per M on 2027-01-01.
Foundry ·Global$0.75CHF 0.604$0.075CHF 0.06$3.75CHF 3.02official
Gemini API pricing page (ai.google.dev/gemini-api/docs/pricing), captured 2026-08-14; page stamped "Last updated 2026-08-13 UTC" and publishes the reversion date inline per price cell, verbatim: "$0.75 through December 31, 2026.$1.50 starting January 1, 2027." Also bills a separate cache-storage dimension at $1.00 per 1M tokens per hour, not modeled by this schema. Batch/Flex bill at exactly 50% of Standard and Priority at exactly 1.8x, in both the current and post-reversion periods — now modeled as service-tier variants below.
Standard (from 2027) begins 1 Jan 2027 UTC
Standard (from 2027)$1.50 / $0.15 / $7.50Batch$0.375 / $0.037 / $1.88Batch (from 2027)$0.75 / $0.075 / $3.75Flex$0.375 / $0.037 / $1.88Flex (from 2027)$0.75 / $0.075 / $3.75Priority$1.35 / $0.135 / $6.75Priority (from 2027)$2.70 / $0.27 / $13.50
Gemini 3.5 Flash
~25% cheaper than 3.1 Pro; beats it on coding/agentic benchmarks.
Foundry ·Global$1.50CHF 1.21$0.15CHF 0.121$9.00CHF 7.25official
Google pricing page. Cached-input rate now published on the same page (captured 2026-07-22).
Confidenceofficial published pagederived reconciled from billing estimate pattern-inferred
Other providers
KimiDeepSeekGLMOpenAI / Azure OpenAIClaudeGrokQwenMistralMiniMaxEmbeddingsCompare all →