ElevenLabs Music generates prompt-driven audio tracks up to five minutes long. Reach for it when your application needs background scores, soundtracks, or royalty-free music from text descriptions.
Modalities
Text → Audio
Price
$0.135 / sec
Speed
slow
Performance
Live production data from real requests on Kyma — not synthetic benchmarks.
Rank
#79
of 87 active models
Tokens served
56
all-time
Pricing
Per second of generated video. Failed jobs are refunded in full.
$0.135 / secWhen to use ElevenLabs Music
Where this model earns its cost — and where it doesn't.
This model converts text prompts into audio, supporting both lyrical and instrumental requests. It accepts a 2000-token context window and produces audio outputs up to five minutes per generation.
On Kyma, the model runs on the premium tier with slower generation speeds. It supports prompt caching, which bills repeated prefixes at 10% of the standard input rate. Responses return exact generation cost in usage.cost and identify the active model via the X-Kyma-Model header, with automatic failover handling routing if a path degrades.
The model does not support reasoning, vision, or structured outputs. It accepts text-only input and returns audio-only output. It is optimized for batch or asynchronous generation rather than real-time streaming.
Generate background tracks
Create royalty-free audio for videos, games, or applications.
Build custom soundtracks
Produce instrumental or lyrical compositions up to five minutes.
Prototype theme music
Iterate on text prompts to test audio concepts before final production.
Automate content audio
Integrate prompt-driven music generation into media pipelines.
Not ideal for: Do not use this model for real-time audio streaming, speech synthesis, or latency-sensitive interactive voice applications.
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit, no card required.
Get API key → - 2
Make your first request
Submit a generation job and poll until it succeeds.
curl https://kymaapi.com/v1/videos/generations \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "elevenlabs-music", "prompt": "A wave breaking on a rocky shore at golden hour, cinematic", "duration": 5 }' # Poll: GET /v1/jobs/{id} until status="succeeded"
FAQ
Common questions about this model.
How much does ElevenLabs Music cost?
How do I use ElevenLabs Music?
Can I pass lyrics to the prompt?
How does prompt caching work for this model?
What happens if the generation endpoint fails?
More models by ElevenLabs
| Model | Context | Input | Output |
|---|---|---|---|
ElevenLabs v3 | — | $0.405 / 1K char | |
ElevenLabs Flash v2.5 | — | $0.2025 / 1K char | |
ElevenLabs Turbo v2.5 | — | $0.2025 / 1K char | |
ElevenLabs Sound Effects | — | $0.027 / call | |
ElevenLabs Multilingual v2 | — | $0.405 / 1K char | |
