MiniMax Voice Design generates synthetic voice profiles from natural-language descriptions without requiring reference audio. Use it when you need custom vocal identities for branding, characters, or personas but lack audio samples.
Modalities
Text → Audio
Price
$4.20 / call
Speed
medium
Performance
Live production data from real requests on Kyma — not synthetic benchmarks.
Rank
#68
of 87 active models
Tokens served
56
all-time
Pricing
Flat per successful generation.
$4.20 / callWhen to use MiniMax Voice Design
Where this model earns its cost — and where it doesn't.
MiniMax Voice Design creates a synthesized voice profile from a text prompt. It returns a voice_id that works with any MiniMax HD or Turbo speech endpoint. The model operates on a per-call basis with a flat charge for each designed voice.
On Kyma, the model runs through the standard OpenAI-compatible base URL using a single API key. Every request includes automatic failover if a serving path degrades, and responses report exact costs in usage.cost alongside the active model in the X-Kyma-Model header. Prompt caching applies to repeated prefixes at a 10% rate, and new accounts include a $0.50 signup credit.
The model accepts text input and outputs audio metadata rather than continuous speech streams. It does not support reasoning, vision, or structured outputs, and operates with a 1,000-token context window. It is optimized for voice design rather than direct text-to-speech generation.
Brand Voice Creation
Generate consistent synthetic voices for corporate branding without hiring voice actors.
Fictional Character Voices
Design unique vocal profiles for game or narrative characters using text prompts.
Persona Voice Generation
Create custom speaker identities from descriptive text when reference audio is unavailable.
Rapid Audio Prototyping
Test different vocal styles during product development using a flat per-call workflow.
Not ideal for: Do not use this model to clone existing voices from audio samples or to generate continuous long-form speech directly.
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit, no card required.
Get API key → - 2
Make your first request
Drop in your key and send a chat completion — fully OpenAI-compatible.
curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "minimax-voice-design", "messages": [ {"role": "user", "content": "Explain prompt caching in one paragraph."} ] }' - 3
Stream responses
Add
"stream": trueto receive tokens as they arrive.curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "minimax-voice-design", "stream": true, "messages": [{"role": "user", "content": "Hello!"}] }'
FAQ
Common questions about this model.
How much does MiniMax Voice Design cost?
How do I use MiniMax Voice Design?
How do I use the generated voice profile?
Does Kyma cache requests for this model?
What happens if the serving path degrades?
More models by MiniMax
See all 13 →| Model | Context | Input | Output |
|---|---|---|---|
MiniMax M3 | 1M | $0.3852 | $1.54 |
MiniMax Music Pro | — | $0.21 / song | |
MiniMax M2.7 | 205K | $0.405 | $1.62 |
MiniMax M2.5 | 197K | $0.3645 | $1.282 |
MiniMax Music | — | $0.045 / song | |
Hailuo 02 (1080p) | — | $0.78 / video | |
Hailuo 02 (768p) | — | $0.42 / video | |
Hailuo 02 (512p) | — | $0.14 / video | |
