384 dimensions from six layers — the most downloaded embedding model there is, and the one most tutorials and starter repos hard-code. Reach for it to match an existing index or to keep a local prototype and a hosted one on the same vectors.
Modalities
Text → Text
Usage
How much this model is actually called here.
Rank
#83
of 98 active models
Tokens served
32.4K
all-time
Platform share
0.0%
of all tokens
Pricing
Pay per token. Cached input bills at this model’s own cached rate, listed below.
+ Estimate your workload− Estimate your workload
Estimated monthly cost
$0.4050
$0.0135 / day on all-MiniLM-L6-v2
Same workload on:
Estimates use list pricing. Actual bills depend on real token counts, and every response includes its exact cost.
How it compares
Against the peers people actually weigh it against.
| Spec | all-MiniLM-L6-v2 | all-MiniLM-L12-v2 | EmbeddingGemma 300M |
|---|---|---|---|
| Input /1M | $0.00675 | $0.00675 | $0.0027 |
| Output /1M | $0.00 | $0.00 | $0.00 |
| Context | 1K | 1K | 2K |
| Tools | No | No | No |
| Reasoning | No | No | No |
| Throughput | 259 tok/s | 10 tok/s | 1211 tok/s |
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit on the free tier, which covers all-MiniLM-L6-v2. No card required.
Get API key → - 2
Make your first request
Submit a generation job and poll until it succeeds.
curl https://kymaapi.com/v1/embeddings \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "all-minilm-l6", "input": ["first document", "second document"] }'
FAQ
Common questions about this model.
What is the context window of all-MiniLM-L6-v2?
How much does the all-MiniLM-L6-v2 API cost?
Does all-MiniLM-L6-v2 support function calling?
How do I use all-MiniLM-L6-v2?
More models by Sentence Transformers
| Model | Context | Input | Output |
|---|---|---|---|
| 1K | $0.00675 | $0.00 |