What are you spending on Gemini?
Google's range is unusually wide: the Flash tier is priced for high-volume work while Pro, Veo, and the image models sit at the frontier. Teams that spend a lot on Gemini usually do it in two very different places at once, so splitting the estimate by model matters more here than most.
Build your basket
Your monthly AI spend
2 models- −25%
- −40%
Enter what you spend today at each provider's direct price. Savings are computed per model from the rates published on this page — models we price the same as the provider show no saving, and are left in rather than quietly dropped to flatter the total.
Your estimate
Current spend
$3,400
per month
Estimated with INFRO
$2,415
per month
Potential savings
$985
29% blended across your basket
Over twelve months
$11,820
same basket, same rates
29% of current spend
An estimate, not a quote. It prices your basket at the reference and INFRO rates published on this page — your real mix of prompt lengths, cached tokens, and image sizes will move it.
What moves this bill more than the vendor does.
A calculator prices a basket. These are the things that make a real invoice differ from one — worth knowing before you treat any estimate, including ours, as a number.
- Flash and Pro are different businesses
- The gap between the Flash and Pro tiers per token is large enough that routing decisions between them typically move a bill more than any vendor change would. Estimate them separately rather than as an average.
- Long context can change the rate
- Google has historically priced very long context requests differently from short ones. If a meaningful share of your traffic runs at hundreds of thousands of tokens, treat any blended estimate — including this one — as a floor.
- Veo is priced per second of output
- As with other video models, cost tracks generated seconds and resolution rather than job count. Estimate by monthly seconds, and remember that regenerations count.
Every Google model we carry
Direct price beside ours, per unit. Rows priced at parity are left in rather than dropped.
| Model | Task | Direct | INFRO | Save |
|---|---|---|---|---|
| Gemini 3.1 Progoogle/gemini-3.1-pro | Multimodal reasoning | $4.50 | $3.38 /1M tok | −25% |
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | Fast agents & coding | $1.50 | $0.90 /1M tok | −40% |
| Nano Banana Progoogle/nano-banana-pro | 2K–4K image generation | $0.18 | $0.09 /image | −50% |
| Nano Banana 2google/nano-banana-2 | Fast image generation | $0.045 | $0.027 /image | −40% |
| Imagen 4 Ultragoogle/imagen-4-ultra | Image generation | $0.06 | $0.046 /image | −23% |
| Veo 3.1google/veo-3.1 | Cinematic video, native audio | $0.40 | $0.24 /second | −40% |
| Gemini 3.1 Flash TTSgoogle/gemini-3.1-flash-tts | Controllable text to speech | $16 | $10 /1M chars | −38% |
USD, illustrative launch rates reconciled on 2026-08-24. Live rates come from GET /v1/models once your key is active.
An estimate is a floor. Want the real number?
Send a usage export and we will price your actual traffic — with the arithmetic shown, and the caveats named.