Alibaba Cloud: qwen3-vl-rerank
qwen3-vl-rerank- Input / Output
- $0.1029 / $0.00
- Context
- 120K
- Modalities
- VisionChat
Getting started
Create API key
Create a key in Console, then use it with every model on the platform.
Send your first request
Copy the example for your language and run it against the endpoint.
RerankingTokenLab endpointPOST/v1/rerankcurl -X POST "https://api.tokenlab.sh/v1/rerank" \ -H "Authorization: Bearer sk-xxx" \ -H "Content-Type: application/json" \ -d '{ "operation": "rerank", "model": "qwen3-vl-rerank", "query": "What is the capital of France?", "documents": [ "Paris is the capital of France.", "Berlin is the capital of Germany." ] }'
Pricing
Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.
Default
Per 1M tokens- Official price
- Input $0.1029 / Text Input $0.1029 / Image Input $0.2647
- TokenLab price
- Input $0.1029 / Text Input $0.1029 / Image Input $0.2647
- Discount
- -
Display Base Rates
Per 1M tokens- Official price
- Output $0.00
- TokenLab price
- Output $0.00
- Discount
- -
| Official pricePer 1M tokens | TokenLab pricePer 1M tokens | Discount | |
|---|---|---|---|
| Default | Input $0.1029 / Text Input $0.1029 / Image Input $0.2647 | Input $0.1029 / Text Input $0.1029 / Image Input $0.2647 | - |
| Display Base Rates | Output $0.00 | Output $0.00 | - |
Usage & activity
Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.
Usage & availability
Last 24 hoursNo data yet
Data is based on aggregate user requests, excluding status checks.
Open in Console
Open Qwen-Rerank in Console with a prompt ready to edit or send.
Help me try qwen3-vl-rerank with rerank at /v1/rerank. Show the result and cost.
Use cases
Best forVision
Reading images, parsing documents, and answering visual questions
Search quality
Rerank retrieval candidates and lift top-result precision for search and RAG.
Side-by-side test
Compare response quality, latency, and price side by side.
Prompt examples
Rerank these search results for the query: best image generation model for product photos.
Send the smallest possible request and show me the result and cost.
Compare this model with a cheaper one in the same category — when is each worth it?
This model has conditional pricing. Monthly token totals alone cannot produce a reliable estimate; use the detailed pricing for the request specification, cache, and applicable time window.
FAQ
How much does Qwen-Rerank cost?
On TokenLab, Qwen-Rerank costs Input $0.1029 / Output $0.00 Per 1M tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
What is Qwen-Rerank best for?
Qwen-Rerank supports Vision. You can open it directly in Create.
How do I test Qwen-Rerank?
Open Qwen-Rerank in Create. A sample for /v1/rerank will be ready to try.
Which endpoint should Qwen-Rerank use?
Use https://api.tokenlab.sh/v1/rerank for Qwen-Rerank. The request example below shows the matching code shape.
Can I test Qwen-Rerank before integrating it?
Yes. Open Console starts a ready draft for Qwen-Rerank and keeps your prompt after sign-in, so you don’t lose context.
Which operations does Qwen-Rerank support?
Qwen-Rerank supports Reranking. Select an operation above to see its endpoint and request example.