Meta's larger Llama 4 tier, above Scout. 1M context, vision, served from two independent providers.
Modalities
Text+Image → Text
Input
$0.27 /1M
Output
$1.08 /1M
Context
1M
Speed
fast
Pricing
Pay per token. Cached input is billed at 10% of the input rate.
+ Estimate your workload− Estimate your workload
Estimated monthly cost
$32.40
$1.08 / day on Llama 4 Maverick
Same workload on:
Estimates use list pricing. Actual bills depend on real token counts — every response includes its exact cost.
How it compares
Against the peers people actually weigh it against.
| Spec | Llama 4 Maverick | Muse Spark 1.1 | Llama 3.3 70B |
|---|---|---|---|
| Input /1M | $0.27 | $1.688 | $0.7965 |
| Output /1M | $1.08 | $5.738 | $1.067 |
| Context | 1M | 1M | 128K |
| Tools | Yes | Yes | Yes |
| Reasoning | No | Yes | Yes |
| Speed | fast | medium | medium |
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit, no card required.
Get API key → - 2
Make your first request
Drop in your key and send a chat completion — fully OpenAI-compatible.
curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "llama-4-maverick", "messages": [ {"role": "user", "content": "Explain prompt caching in one paragraph."} ] }' - 3
Stream responses
Add
"stream": trueto receive tokens as they arrive.curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "llama-4-maverick", "stream": true, "messages": [{"role": "user", "content": "Hello!"}] }'
FAQ
Common questions about this model.
What is the context window of Llama 4 Maverick?
How much does the Llama 4 Maverick API cost?
Does Llama 4 Maverick support function calling?
How do I use Llama 4 Maverick?
More models by Meta
| Model | Context | Input | Output |
|---|---|---|---|
Muse Spark 1.1 | 1M | $1.688 | $5.738 |
Llama 3.3 70B | 128K | $0.7965 | $1.067 |
