Build next-gen apps with qwen-flash
Pricing
Input tokens <= 125K
Per 1M Tokens- Official Price
- Input $0.0221 / Output $0.2206 / Cache Read $0.004412 / Cache Write $0.0276
- TokenLab Price
- Input $0.0221 / Output $0.2206 / Cache Read $0.004412 / Cache Write $0.0276
- Discount
- -
Input tokens <= 250K
Per 1M Tokens- Official Price
- Input $0.0882 / Output $0.8824 / Cache Read $0.0176 / Cache Write $0.1103
- TokenLab Price
- Input $0.0882 / Output $0.8824 / Cache Read $0.0176 / Cache Write $0.1103
- Discount
- -
Input tokens <= 1M
Per 1M Tokens- Official Price
- Input $0.1765 / Output $1.76 / Cache Read $0.0353 / Cache Write $0.2206
- TokenLab Price
- Input $0.1765 / Output $1.76 / Cache Read $0.0353 / Cache Write $0.2206
- Discount
- -
| Official PricePer 1M Tokens | TokenLab PricePer 1M Tokens | Discount | |
|---|---|---|---|
| Input tokens <= 125K | Input $0.0221 / Output $0.2206 / Cache Read $0.004412 / Cache Write $0.0276 | Input $0.0221 / Output $0.2206 / Cache Read $0.004412 / Cache Write $0.0276 | - |
| Input tokens <= 250K | Input $0.0882 / Output $0.8824 / Cache Read $0.0176 / Cache Write $0.1103 | Input $0.0882 / Output $0.8824 / Cache Read $0.0176 / Cache Write $0.1103 | - |
| Input tokens <= 1M | Input $0.1765 / Output $1.76 / Cache Read $0.0353 / Cache Write $0.2206 | Input $0.1765 / Output $1.76 / Cache Read $0.0353 / Cache Write $0.2206 | - |
| Cache Read | $0.004412 | $0.004412 | - |
| Cache Write | $0.0276 | $0.0276 | - |
One-click test
Sign in once and Web Agent keeps this model, prompt, and request preset for you.
Test qwen-flash in Web Agent with a short request at /v1/messages, then show request body, latency, and response.
API workbench
The default route for production. The code sample below uses this endpoint with the format you pick.
curl https://api.tokenlab.sh/v1/messages \
-H "Content-Type: application/json" \
-H "x-api-key: sk-xxx" \
-H "anthropic-version: 2023-06-01" \
-d '{
"model": "qwen-flash",
"max_tokens": 1024,
"messages": [
{"role": "user", "content": "Hello!"}
]
}'Use cases
Agents and tools
Drive reasoning, support triage, tool calls, and multi-step task flows.
Developer workflows
Generate, review, or debug code without rewiring your stack.
Knowledge assistants
Ship chat, search, and retrieval with predictable cost and behavior.
Side-by-side test
Compare real response quality, latency, and price for production defaults.
Prompt examples
Write a concise support reply and list the assumptions behind it.
Review this API design and call out the top three integration risks.
Turn a long changelog into release notes a non-engineer would read.
Cost Calculator
FAQ
How much does qwen-flash cost?
On TokenLab, qwen-flash costs Input $0.0221 / Output $0.2206 Per 1M Tokens. The pricing table above shows the full breakdown.
What is qwen-flash best for?
qwen-flash is a strong fit for Prompt Cache. You can call it through TokenLab with one API key.
How do I call the qwen-flash API?
Get a TokenLab API key, then send your request to https://api.tokenlab.sh/v1/messages. The API workbench above has a recommended endpoint and copy-ready code.
Which endpoint should qwen-flash use?
Use https://api.tokenlab.sh/v1/messages as the default for qwen-flash. If a provider-native format is supported, the API workbench shows that endpoint too.
Can I test qwen-flash before integrating it?
Yes. Try in Web Agent opens a ready test for qwen-flash and keeps your prompt after sign-in, so you don’t lose context.
Which operations does qwen-flash support?
qwen-flash supports Text to text. Pick an operation in the API workbench above to see the matching endpoint and code sample.