Chat Models

AI models for chat, Q&A, reasoning, and text generation

Chat · Moonshot

kimi-k3
Available
JSON modePrompt CacheReasoningTool useVisionCached inputless
TokenLab price— Discount
—
Official price
Input$3.00Output$15.00/1M
kimi-k2.5
DeprecatedAvailable
JSON modePrompt CacheTool useVisionCached inputless
TokenLab price— Discount
—
Official price
Input$0.60Output$3.00/1M
kimi-k2.6
Available
JSON modePrompt CacheTool useVisionCached inputless
TokenLab price— Discount
—
Official price
Input$0.95Output$4.00/1M
kimi-k2.7-code
Available
JSON modePrompt CacheTool useVisionCached inputless
TokenLab price/ Official price— Discount
—
kimi-k2.7-code-highspeed
Available
JSON modePrompt CacheTool useVisionCached inputless
TokenLab price— Discount
—
Official price
Input$1.90Output$8.00/1M

All model series

1 series

Choosing the right chat model

The right model balances task fit, output quality, latency, and price.

Selection signals

  • Match the model to the input and output you actually need.
  • Compare models that use the same pricing unit.
  • A small test reveals quality and latency differences that a catalog cannot.

FAQ

What makes a good chat model?

A strong fit matches the task, quality bar, latency target, and budget. A small evaluation set makes quality and reliability differences easier to see.

Can I switch models later?

Yes. Keep the public model ID and request format explicit, and compare alternatives on the same tasks.