384-dimension embeddings — the smallest vector on the shelf, and the default of the sentence-transformers library, so an enormous amount of existing code expects exactly this shape. Twelve layers where the L6 has six: better recall, still tiny.
Modalities
Text → Text
Usage
How much this model is actually called here.
Rank
#85
of 98 active models
Tokens served
15.9K
all-time
Platform share
0.0%
of all tokens
Pricing
Pay per token. Cached input bills at this model’s own cached rate, listed below.
+ Estimate your workload− Estimate your workload
Estimated monthly cost
$0.4050
$0.0135 / day on all-MiniLM-L12-v2
Same workload on:
Estimates use list pricing. Actual bills depend on real token counts, and every response includes its exact cost.
How it compares
Against the peers people actually weigh it against.
| Spec | all-MiniLM-L12-v2 | all-MiniLM-L6-v2 | EmbeddingGemma 300M |
|---|---|---|---|
| Input /1M | $0.00675 | $0.00675 | $0.0027 |
| Output /1M | $0.00 | $0.00 | $0.00 |
| Context | 1K | 1K | 2K |
| Tools | No | No | No |
| Reasoning | No | No | No |
| Throughput | 10 tok/s | 259 tok/s | 1211 tok/s |
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit on the free tier, which covers all-MiniLM-L12-v2. No card required.
Get API key → - 2
Make your first request
Submit a generation job and poll until it succeeds.
curl https://kymaapi.com/v1/embeddings \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "all-minilm-l12", "input": ["first document", "second document"] }'
FAQ
Common questions about this model.
What is the context window of all-MiniLM-L12-v2?
How much does the all-MiniLM-L12-v2 API cost?
Does all-MiniLM-L12-v2 support function calling?
How do I use all-MiniLM-L12-v2?
More models by Sentence Transformers
| Model | Context | Input | Output |
|---|---|---|---|
| 1K | $0.00675 | $0.00 |