Skip to content
INFRO

Google · Audio

Gemini 3.1 Flash TTS API pricing

google/gemini-3.1-flash-tts

Controllable text to speech. Gemini 3.1 Flash TTS by Google, called through one INFRO key alongside every other model you run. We price it 38% below Google's direct rate — the same model, the same weights, reached through a different integration.

Pricing

Per 1M charsUSD · updated 2026-08-24

Google direct

$16

per 1M chars

INFRO

$10

per 1M chars

You save

38%

on every unit

Illustrative launch rates, reconciled against published provider pricing on 2026-08-24. Live rates come from GET /v1/models once your key is active.

Calling Gemini 3.1 Flash TTS

POST /v1/audio/speechAt launchCommitted to the first release. In build now, not usable yet.
import os
from infro import Infro

client = Infro(api_key=os.environ["INFRO_API_KEY"])

audio = client.audio.speech.create(
    model="google/gemini-3.1-flash-tts",
    input="Every AI model, one API.",
)

What this modality supports

  • Text to speech
  • Music generation
  • Transcription
  • Diarization

Text-to-speech with streaming playback, full-track music generation, and transcription with timestamps and diarization. Billed per character generated, per track, and per hour transcribed.

Same modality, nearest in price. All of them share the request shape above, so comparing them is a one-word change.

Work out what Gemini 3.1 Flash TTS would cost you

Put your current spend on this model into the calculator and see the same workload priced here — or have us read your real usage.