AvailableZhipu AIChatCache 75% off
glm-5.2

Build next-gen apps with GLM

API Code Example

Pricing

Official Price

Input$1.18
Output$4.12
Per 1M Tokens

TokenLab Price

Input$1.18
Output$4.12
Per 1M Tokens
价格详情
每 1M Tokens
支持
文本到文本
定价层级
规格
输入 官方 $1.18/1M tokens · 平台 $1.18/1M tokens 输出 官方 $4.12/1M tokens · 平台 $4.12/1M tokens 缓存读取 官方 $0.2941/1M tokens · 平台 $0.2941/1M tokens 缓存创建 官方 $0.00/1M tokens · 平台 $0.00/1M tokens 文本 输入 官方 $1.18/1M tokens · 平台 $1.18/1M tokens 文本 输出 官方 $4.12/1M tokens · 平台 $4.12/1M tokens

One-click test

Sign in once and Web Agent keeps this model, prompt, and request preset for you.

Test glm-5.2 in Web Agent with a short request at /v1/messages, then show request body, latency, and response.

API workbench

The default route for production. The code sample below uses this endpoint with the format you pick.

ChatOpenAI Compatible
POST/v1/messages
API Format:
curl https://api.tokenlab.sh/v1/messages \
  -H "Content-Type: application/json" \
  -H "x-api-key: sk-xxx" \
  -H "anthropic-version: 2023-06-01" \
  -d '{
    "model": "glm-5.2",
    "max_tokens": 1024,
    "messages": [
      {"role": "user", "content": "Hello!"}
    ]
  }'

Use cases

01

Agents and tools

Drive reasoning, support triage, tool calls, and multi-step task flows.

02

Developer workflows

Generate, review, or debug code without rewiring your stack.

03

Knowledge assistants

Ship chat, search, and retrieval with predictable cost and behavior.

04

Side-by-side test

Compare real response quality, latency, and price for production defaults.

Prompt examples

Write a concise support reply and list the assumptions behind it.

Review this API design and call out the top three integration risks.

Turn a long changelog into release notes a non-engineer would read.

此模型采用条件计费。仅凭每月 token 总量无法可靠估算费用;请结合请求规格、缓存及适用时段查看详细价格。

FAQ

How much does GLM cost?

On TokenLab, GLM costs Input $1.18 / Output $4.12 Per 1M Tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What is GLM best for?

GLM is a strong fit for JSON Mode, Prompt Cache, Tool Use. You can call it through TokenLab with one API key.

How do I call the GLM API?

Get a TokenLab API key, then send your request to https://api.tokenlab.sh/v1/messages. The API workbench above has a recommended endpoint and copy-ready code.

Which endpoint should GLM use?

Use https://api.tokenlab.sh/v1/messages as the default for GLM. If a provider-native format is supported, the API workbench shows that endpoint too.

Can I test GLM before integrating it?

Yes. Try in Web Agent opens a ready test for GLM and keeps your prompt after sign-in, so you don’t lose context.

Related Models