Skip to content
INFRO

What are you spending on Gemini?

Google's range is unusually wide: the Flash tier is priced for high-volume work while Pro, Veo, and the image models sit at the frontier. Teams that spend a lot on Gemini usually do it in two very different places at once, so splitting the estimate by model matters more here than most.

Build your basket

Cost auditMonthly · USD

Your monthly AI spend

2 models
  • −25%
  • −40%

Enter what you spend today at each provider's direct price. Savings are computed per model from the rates published on this page — models we price the same as the provider show no saving, and are left in rather than quietly dropped to flatter the total.

Your estimate

Current spend

$3,400

per month

Estimated with INFRO

$2,415

per month

Potential savings

$985

29% blended across your basket

Over twelve months

$11,820

same basket, same rates

29% of current spend

Get a free cost auditOr request early access →

An estimate, not a quote. It prices your basket at the reference and INFRO rates published on this page — your real mix of prompt lengths, cached tokens, and image sizes will move it.

What moves this bill more than the vendor does.

A calculator prices a basket. These are the things that make a real invoice differ from one — worth knowing before you treat any estimate, including ours, as a number.

Flash and Pro are different businesses
The gap between the Flash and Pro tiers per token is large enough that routing decisions between them typically move a bill more than any vendor change would. Estimate them separately rather than as an average.
Long context can change the rate
Google has historically priced very long context requests differently from short ones. If a meaningful share of your traffic runs at hundreds of thousands of tokens, treat any blended estimate — including this one — as a floor.
Veo is priced per second of output
As with other video models, cost tracks generated seconds and resolution rather than job count. Estimate by monthly seconds, and remember that regenerations count.

Every Google model we carry

Direct price beside ours, per unit. Rows priced at parity are left in rather than dropped.

ModelTaskDirectINFROSave
Gemini 3.1 Progoogle/gemini-3.1-proMultimodal reasoning$4.50$3.38 /1M tok25%
Gemini 3.7 Flashgoogle/gemini-3.7-flashFast agents & coding$1.50$0.90 /1M tok40%
Nano Banana Progoogle/nano-banana-pro2K–4K image generation$0.18$0.09 /image50%
Nano Banana 2google/nano-banana-2Fast image generation$0.045$0.027 /image40%
Imagen 4 Ultragoogle/imagen-4-ultraImage generation$0.06$0.046 /image23%
Veo 3.1google/veo-3.1Cinematic video, native audio$0.40$0.24 /second40%
Gemini 3.1 Flash TTSgoogle/gemini-3.1-flash-ttsControllable text to speech$16$10 /1M chars38%

USD, illustrative launch rates reconciled on 2026-08-24. Live rates come from GET /v1/models once your key is active.

An estimate is a floor. Want the real number?

Send a usage export and we will price your actual traffic — with the arithmetic shown, and the caveats named.