Alibaba

Alibaba

Qwen3 Reranker 8B

Scores how well each document answers a query, so a retriever's top 50 can be cut to the 5 worth sending to a model. Sits between search and synthesis: embeddings decide what to fetch, this decides what survives. A 40K window means whole documents can be judged without chunking.

Modalities

Text → Text

Usage

How much this model is actually called here.

Rank

#86

of 98 active models

Tokens served

11.7K

all-time

Platform share

0.0%

of all tokens

Pricing

Pay per token. Cached input bills at this model’s own cached rate, listed below.

$0.135 /1M input$0.00 /1M output
Hobby10 req/day · 2K in / 500 out
~$0.08/mo
Production1,000 req/day · 2K in / 500 out
~$8.10/mo
Scale20,000 req/day · 2K in / 500 out
~$162/mo
+ Estimate your workload
1,000
2,000
500

Estimated monthly cost

$8.10

$0.2700 / day on Qwen3 Reranker 8B

Same workload on:

Qwen 3.8 27B$94.77+1070%
Qwen 3.7 Flash$6.23-23%

Estimates use list pricing. Actual bills depend on real token counts, and every response includes its exact cost.

How it compares

Against the peers people actually weigh it against.

SpecQwen3 Reranker 8BQwen 3.8 27BQwen 3.7 Flash
Input /1M$0.135$0.567$0.0498
Output /1M$0.00$4.05$0.2164
Context41K1M1M
ToolsNoYesYes
ReasoningNoYesYes
Throughput46 tok/s118 tok/s

Quick start

Up and running in under two minutes.

  1. 1

    Create an API key

    Sign up and grab a key from the dashboard — $0.50 free credit on the free tier, which covers Qwen3 Reranker 8B. No card required.

    Get API key →
  2. 2

    Make your first request

    Submit a generation job and poll until it succeeds.

    curl https://kymaapi.com/v1/rerank \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "qwen3-reranker-8b",
        "query": "what is prompt caching?",
        "documents": ["...", "..."]
      }'

FAQ

Common questions about this model.

What is the context window of Qwen3 Reranker 8B?

Qwen3 Reranker 8B has a 41K-token context window — roughly 60 pages of text in a single request.

How much does the Qwen3 Reranker 8B API cost?

$0.135 per 1M input tokens and $0.00 per 1M output tokens. No subscription; you pay only for what you use.

Does Qwen3 Reranker 8B support function calling?

No — Qwen3 Reranker 8B does not support tool calling. For agent workloads, choose a tools-enabled model from the catalog.

How do I use Qwen3 Reranker 8B?

Kyma is OpenAI-compatible: point your SDK's base URL at https://kymaapi.com/v1, use your Kyma API key, and set the model to qwen3-reranker-8b. Signing up is free and includes $0.50 of credit on the free tier, which covers this model — no card required.

Start with $0.50 free credit on the free tier — no card required.Create account →

More models by Alibaba

See all 14
ModelContextInputOutput
AlibabaQwen 3.8 Flash1M$0.203$0.635
AlibabaQwen 3.8 27B1M$0.567$4.05
AlibabaQwen 3.8 Max1M$2.2275$6.684
AlibabaQwen 3.7 Flash1M$0.0498$0.2164
AlibabaQwen 3.7 Plus1M$0.4431$1.773
AlibabaQwen 3.7 Max1M$2.304$6.909
AlibabaQwen 3.6 Plus1M$0.4911$2.947
AlibabaQwen 3 Coder131K$0.334$1.519