BAAI

BAAI

BGE Base EN v1.5

768-dimension English embeddings — half the vector of bge-large-en at the same per-million rate, so an index built on it is half the storage and half the comparison cost. Retrieval quality gives up little on short chunks, which is what the 512-token window enforces anyway.

Modalities

Text → Text

Input

$0.00675 /1M

Output

$0.00 /1M

Cached input

$0.00068 /1M90% off

Context

1K

Speed

fast

Pricing

Pay per token. Cached input is billed at 10% of the input rate.

$0.00675 /1M input$0.00 /1M output
Hobby10 req/day · 2K in / 500 out
~$0.00/mo
Production1,000 req/day · 2K in / 500 out
~$0.40/mo
Scale20,000 req/day · 2K in / 500 out
~$8.10/mo
+ Estimate your workload
1,000
2,000
500
30%

Estimated monthly cost

$0.2956

$0.0099 / day on BGE Base EN v1.5

Same workload on:

BGE-M3$0.8100+174%
BGE Large EN v1.5$0.8100+174%

Estimates use list pricing with cached input billed at the 90%-discount rate. Actual bills depend on real token counts — every response includes its exact cost.

How it compares

Against the peers people actually weigh it against.

SpecBGE Base EN v1.5BGE-M3BGE Large EN v1.5
Input /1M$0.00675$0.0135$0.0135
Output /1M$0.00$0.00$0.00
Context1K8K1K
ToolsYesYesYes
ReasoningNoNoNo
Speedfastfastfast

Quick start

Up and running in under two minutes.

  1. 1

    Create an API key

    Sign up and grab a key from the dashboard — $0.50 free credit, no card required.

    Get API key →
  2. 2

    Make your first request

    Drop in your key and send a chat completion — fully OpenAI-compatible.

    curl https://kymaapi.com/v1/chat/completions \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "bge-base-en",
        "messages": [
          {"role": "user", "content": "Explain prompt caching in one paragraph."}
        ]
      }'
  3. 3

    Stream responses

    Add "stream": true to receive tokens as they arrive.

    curl https://kymaapi.com/v1/chat/completions \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "bge-base-en",
        "stream": true,
        "messages": [{"role": "user", "content": "Hello!"}]
      }'

FAQ

Common questions about this model.

What is the context window of BGE Base EN v1.5?

BGE Base EN v1.5 has a 1K-token context window — roughly 1 pages of text in a single request.

How much does the BGE Base EN v1.5 API cost?

$0.00675 per 1M input tokens and $0.00 per 1M output tokens, with cached input at $0.00068/1M — a 90% discount on repeated prompt prefixes. No subscription; you pay only for what you use.

Does BGE Base EN v1.5 support function calling?

Yes — BGE Base EN v1.5 supports tool/function calling and structured outputs (JSON mode), so it works with agent frameworks out of the box.

How do I use BGE Base EN v1.5?

Kyma is OpenAI-compatible: point your SDK's base URL at https://kymaapi.com/v1, use your Kyma API key, and set the model to bge-base-en. Signing up is free and includes $0.50 of credit — no card required.

Start with $0.50 free credit — no card required.Create account →

More models by BAAI

ModelContextInputOutput
BAAIBGE-M38K$0.0135$0.00
BAAIBGE Large EN v1.51K$0.0135$0.00