50K
text
audio
Not listed
About this model
MiniMax Speech 2.8 Turbo is the speed-optimized variant of MiniMax's flagship text-to-speech model. Delivering sub-250 millisecond latency, it provides fast, natural, and expressive voice synthesis ideal for real-time applications like conversational AI, voice agents, and interactive experiences. It retains the core advanced capabilities of the HD tier—including 40+ language support, rapid 10-second voice cloning, emotion control, and native sound tags—while prioritizing high-throughput generation and cost-efficiency.
Input modalities
text
Accepted as model input
Providers
Available routing options for this model through OneInfer.
Text to speech
$60.000 / 1M characters
Routing
OneInfer optimized
Pricing
Current OneInfer pricing for this model.
| Usage | Price |
|---|---|
| Text to speechTTS | $60.000 / 1M characters |
Performance
Published evaluation results associated with this model.
API example
curl -X POST https://api.oneinfer.ai/v1/ula/generate-audio \
-H "Authorization: Bearer YOUR_JWT_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"provider": "minimax",
"model": "MiniMaxAI/speech-2.8-turbo",
"prompt": "OneInfer makes AI inference simple.",
"stream": false,
"voice_id": "English_expressive_narrator",
"format": "mp3"
}'