MiniMax · Audio
Speech 2.6 API pricing
minimax/speech-2.6
Text to speech. Speech 2.6 by MiniMax, called through one INFRO key alongside every other model you run. We price it 40% below MiniMax's direct rate — the same model, the same weights, reached through a different integration.
Pricing
MiniMax direct
$35
per 1M chars
INFRO
$21
per 1M chars
You save
40%
on every unit
Illustrative launch rates, reconciled against published provider pricing on 2026-08-24. Live rates come from GET /v1/models once your key is active.
Calling Speech 2.6
POST /v1/audio/speechAt launch — Committed to the first release. In build now, not usable yet.import os
from infro import Infro
client = Infro(api_key=os.environ["INFRO_API_KEY"])
audio = client.audio.speech.create(
model="minimax/speech-2.6",
input="Every AI model, one API.",
)What this modality supports
- Text to speech
- Music generation
- Transcription
- Diarization
Text-to-speech with streaming playback, full-track music generation, and transcription with timestamps and diarization. Billed per character generated, per track, and per hour transcribed.
Alternatives for text to speech
Same modality, nearest in price. All of them share the request shape above, so comparing them is a one-word change.
Work out what Speech 2.6 would cost you
Put your current spend on this model into the calculator and see the same workload priced here — or have us read your real usage.