MiniMax

MiniMax

MiniMax Voice Design

MiniMax Voice Design generates synthetic voice profiles from natural-language descriptions without requiring reference audio. Use it when you need custom vocal identities for branding, characters, or personas but lack audio samples.

Modalities

Text → Audio

Price

$4.20 / call

Speed

medium

Performance

Live production data from real requests on Kyma — not synthetic benchmarks.

Rank

#76

of 87 active models

Tokens served

56

all-time

Total requests2
Platform share0.0%

Pricing

Flat per successful generation.

$4.20 / call

When to use MiniMax Voice Design

Updated 2026-07-31

Where this model earns its cost — and where it doesn't.

MiniMax Voice Design creates a synthesized voice profile from a text prompt. It returns a voice_id that works with any MiniMax HD or Turbo speech endpoint. The model operates on a per-call basis with a flat charge for each designed voice.

On Kyma, the model runs through the standard OpenAI-compatible base URL using a single API key. Every request includes automatic failover if a serving path degrades, and responses report exact costs in usage.cost alongside the active model in the X-Kyma-Model header. Prompt caching applies to repeated prefixes at a 10% rate, and new accounts include a $0.50 signup credit.

The model accepts text input and outputs audio metadata rather than continuous speech streams. It does not support reasoning, vision, or structured outputs, and operates with a 1,000-token context window. It is optimized for voice design rather than direct text-to-speech generation.

Brand Voice Creation

Generate consistent synthetic voices for corporate branding without hiring voice actors.

Fictional Character Voices

Design unique vocal profiles for game or narrative characters using text prompts.

Persona Voice Generation

Create custom speaker identities from descriptive text when reference audio is unavailable.

Rapid Audio Prototyping

Test different vocal styles during product development using a flat per-call workflow.

Not ideal for: Do not use this model to clone existing voices from audio samples or to generate continuous long-form speech directly.

Quick start

Up and running in under two minutes.

  1. 1

    Create an API key

    Sign up and grab a key from the dashboard — $0.50 free credit, no card required.

    Get API key →
  2. 2

    Make your first request

    Drop in your key and send a chat completion — fully OpenAI-compatible.

    curl https://kymaapi.com/v1/chat/completions \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "minimax-voice-design",
        "messages": [
          {"role": "user", "content": "Explain prompt caching in one paragraph."}
        ]
      }'
  3. 3

    Stream responses

    Add "stream": true to receive tokens as they arrive.

    curl https://kymaapi.com/v1/chat/completions \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "minimax-voice-design",
        "stream": true,
        "messages": [{"role": "user", "content": "Hello!"}]
      }'

FAQ

Common questions about this model.

How much does MiniMax Voice Design cost?

$4.20 per call. Flat per successful generation.

How do I use MiniMax Voice Design?

Kyma is OpenAI-compatible: point your SDK's base URL at https://kymaapi.com/v1, use your Kyma API key, and set the model to minimax-voice-design. Signing up is free and includes $0.50 of credit — no card required.

How do I use the generated voice profile?

The model returns a voice_id that you pass to the /v1/audio/speech endpoint with any MiniMax HD or Turbo SKU.

Does Kyma cache requests for this model?

Yes, prompt caching is supported and bills repeated prompt prefixes at 10% of the standard input rate.

What happens if the serving path degrades?

Kyma handles automatic failover, so degraded paths are rerouted transparently without requiring client-side retries.

Start with $0.50 free credit — no card required.Create account →

More models by MiniMax

See all 13
ModelContextInputOutput
MiniMaxMiniMax M31M$0.3852$1.54
MiniMaxMiniMax Music Pro$0.21 / song
MiniMaxMiniMax M2.7205K$0.405$1.62
MiniMaxMiniMax M2.5197K$0.3645$1.282
MiniMaxMiniMax Music$0.045 / song
MiniMaxHailuo 02 (1080p)$0.78 / video
MiniMaxHailuo 02 (768p)$0.42 / video
MiniMaxHailuo 02 (512p)$0.14 / video