AvailableGoogleChatCache 90% off
gemini-3.8-flashNew

Build next-gen apps with gemini-3.8-flash

API Code Example

Pricing

Official Price

Input$0.75
Output$3.75
Per 1M Tokens

TokenLab Price

Input$0.375
Output$1.88
Per 1M Tokens
价格详情
每 1M Tokens
定价层级
规格
输入 官方 $0.75/1M tokens · 平台 $0.375/1M tokens 输出 官方 $3.75/1M tokens · 平台 $1.88/1M tokens 缓存读取 官方 $0.075/1M tokens · 平台 $0.0375/1M tokens

One-click test

Sign in once and Web Agent keeps this model, prompt, and request preset for you.

Test gemini-3.8-flash in Web Agent with a short request at /v1beta/models/gemini-3.8-flash:generateContent, then show request body, latency, and response.

API workbench

The default route for production. The code sample below uses this endpoint with the format you pick.

ChatOpenAI Compatible
POST/v1beta/models/gemini-3.8-flash:generateContent
API Format:
curl "https://api.tokenlab.sh/v1beta/models/gemini-3.8-flash:generateContent" \
  -H "Content-Type: application/json" \
  -H "x-goog-api-key: sk-xxx" \
  -d '{
    "contents": [
      {"role": "user", "parts": [{"text": "Hello!"}]}
    ]
  }'

Use cases

Best for

Vision

Reading images, parsing documents, and answering visual questions

01

Agents and tools

Drive reasoning, support triage, tool calls, and multi-step task flows.

02

Developer workflows

Generate, review, or debug code without rewiring your stack.

03

Knowledge assistants

Ship chat, search, and retrieval with predictable cost and behavior.

04

Side-by-side test

Compare real response quality, latency, and price for production defaults.

Prompt examples

Write a concise support reply and list the assumptions behind it.

Review this API design and call out the top three integration risks.

Turn a long changelog into release notes a non-engineer would read.

此模型采用条件计费。仅凭每月 token 总量无法可靠估算费用;请结合请求规格、缓存及适用时段查看详细价格。

FAQ

How much does gemini-3.8-flash cost?

On TokenLab, gemini-3.8-flash costs Input $0.375 / Output $1.88 Per 1M Tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What is gemini-3.8-flash best for?

gemini-3.8-flash is a strong fit for JSON Mode, Prompt Cache, Tool Use. You can call it through TokenLab with one API key.

How do I call the gemini-3.8-flash API?

Get a TokenLab API key, then send your request to https://api.tokenlab.sh/v1beta/models/gemini-3.8-flash:generateContent. The API workbench above has a recommended endpoint and copy-ready code.

Which endpoint should gemini-3.8-flash use?

Use https://api.tokenlab.sh/v1beta/models/gemini-3.8-flash:generateContent as the default for gemini-3.8-flash. If a provider-native format is supported, the API workbench shows that endpoint too.

Can I test gemini-3.8-flash before integrating it?

Yes. Try in Web Agent opens a ready test for gemini-3.8-flash and keeps your prompt after sign-in, so you don’t lose context.

Related Models