Chat Models

AI models for chat, Q&A, reasoning, and text generation

Chat · Alibaba Cloud

qwen-vl-ocr
Available
Vision
TokenLab price— Discount
—
Official price
Input$0.0441Output$0.0735/1M
tongyi-xiaomi-analysis-flash
Available
TokenLab price— Discount
—
Official price
Input$0.0294Output$0.0588/1M
tongyi-xiaomi-analysis-pro
Available
TokenLab price— Discount
—
Official price
Input$0.1471Output$0.3971/1M
qwen-mt-flash
Available
TokenLab price— Discount
—
Official price
Input$0.1029Output$0.2868/1M
qwen-mt-lite
Available
TokenLab price— Discount
—
Official price
Input$0.0882Output$0.2353/1M
qwen-mt-plus
Available
TokenLab price— Discount
—
Official price
Input$0.2647Output$0.7941/1M
tongyi-intent-detect-v3
Available
TokenLab price— Discount
—
Official price
Input$0.0588Output$0.1471/1M
2 / 2

All model series

5 series

Choosing the right chat model

The right model balances task fit, output quality, latency, and price.

Selection signals

  • Match the model to the input and output you actually need.
  • Compare models that use the same pricing unit.
  • A small test reveals quality and latency differences that a catalog cannot.

FAQ

What makes a good chat model?

A strong fit matches the task, quality bar, latency target, and budget. A small evaluation set makes quality and reliability differences easier to see.

Can I switch models later?

Yes. Keep the public model ID and request format explicit, and compare alternatives on the same tasks.