768-dimension general text embeddings from Alibaba's Tongyi lab, trained on a broader mixture than BGE and often a point or two ahead of it on out-of-domain retrieval. Same size and same price as bge-base-en, so the choice is which one your corpus likes.
Modalities
Text → Text
Input
$0.00675 /1M
Output
$0.00 /1M
Cached input
$0.00068 /1M90% off
Context
1K
Speed
fast
Pricing
Pay per token. Cached input is billed at 10% of the input rate.
+ Estimate your workload− Estimate your workload
Estimated monthly cost
$0.2956
$0.0099 / day on GTE Base
Same workload on:
Estimates use list pricing with cached input billed at the 90%-discount rate. Actual bills depend on real token counts — every response includes its exact cost.
How it compares
Against the peers people actually weigh it against.
| Spec | GTE Base | Qwen 3.7 Flash | Qwen 3.7 Max |
|---|---|---|---|
| Input /1M | $0.00675 | $0.0458 | $2.56 |
| Output /1M | $0.00 | $0.1987 | $7.676 |
| Context | 1K | 1M | 1M |
| Tools | Yes | Yes | Yes |
| Reasoning | No | Yes | Yes |
| Speed | fast | fast | medium |
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit, no card required.
Get API key → - 2
Make your first request
Drop in your key and send a chat completion — fully OpenAI-compatible.
curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "gte-base", "messages": [ {"role": "user", "content": "Explain prompt caching in one paragraph."} ] }' - 3
Stream responses
Add
"stream": trueto receive tokens as they arrive.curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "gte-base", "stream": true, "messages": [{"role": "user", "content": "Hello!"}] }'
FAQ
Common questions about this model.
What is the context window of GTE Base?
How much does the GTE Base API cost?
Does GTE Base support function calling?
How do I use GTE Base?
More models by Alibaba
See all 11 →| Model | Context | Input | Output |
|---|---|---|---|
Qwen 3.7 Flash | 1M | $0.0458 | $0.1987 |
Qwen 3.7 Plus | 1M | $0.4482 | $1.793 |
Qwen 3.7 Max | 1M | $2.56 | $7.676 |
Qwen 3.6 Plus | 131K | $0.4388 | $2.633 |
Qwen 3 Coder | 131K | $0.297 | $1.35 |
Qwen3 Reranker 8B | 41K | $0.135 | $0.00 |
Qwen3 Embedding 4B | 33K | $0.027 | $0.00 |
Qwen3 Embedding 0.6B | 33K | $0.0135 | $0.00 |
