AvailableGoogleChatCache 90% off
gemini-3.5-flash

Build next-gen apps with gemini-3.5-flash

API Code Example

Pricing

Official Price

Input$1.50
Output$9.00
Per 1M Tokens

TokenLab Price

Input$0.75
Output$4.50
Per 1M Tokens
价格详情
每 1M Tokens
定价层级
规格
输入 官方 $1.50/1M tokens · 平台 $0.75/1M tokens 输出 官方 $9.00/1M tokens · 平台 $4.50/1M tokens 缓存读取 官方 $0.15/1M tokens · 平台 $0.075/1M tokens

One-click test

Sign in once and Web Agent keeps this model, prompt, and request preset for you.

Test gemini-3.5-flash in Web Agent with a short request at /v1beta/models/gemini-3.5-flash:generateContent, then show request body, latency, and response.

API workbench

The default route for production. The code sample below uses this endpoint with the format you pick.

ChatOpenAI Compatible
POST/v1beta/models/gemini-3.5-flash:generateContent
API Format:
curl "https://api.tokenlab.sh/v1beta/models/gemini-3.5-flash:generateContent" \
  -H "Content-Type: application/json" \
  -H "x-goog-api-key: sk-xxx" \
  -d '{
    "contents": [
      {"role": "user", "parts": [{"text": "Hello!"}]}
    ]
  }'

Use cases

Best for

Vision

Reading images, parsing documents, and answering visual questions

01

Agents and tools

Drive reasoning, support triage, tool calls, and multi-step task flows.

02

Developer workflows

Generate, review, or debug code without rewiring your stack.

03

Knowledge assistants

Ship chat, search, and retrieval with predictable cost and behavior.

04

Side-by-side test

Compare real response quality, latency, and price for production defaults.

Prompt examples

Write a concise support reply and list the assumptions behind it.

Review this API design and call out the top three integration risks.

Turn a long changelog into release notes a non-engineer would read.

此模型采用条件计费。仅凭每月 token 总量无法可靠估算费用;请结合请求规格、缓存及适用时段查看详细价格。

FAQ

How much does gemini-3.5-flash cost?

On TokenLab, gemini-3.5-flash costs Input $0.75 / Output $4.50 Per 1M Tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What is gemini-3.5-flash best for?

gemini-3.5-flash is a strong fit for JSON Mode, Prompt Cache, Tool Use. You can call it through TokenLab with one API key.

How do I call the gemini-3.5-flash API?

Get a TokenLab API key, then send your request to https://api.tokenlab.sh/v1beta/models/gemini-3.5-flash:generateContent. The API workbench above has a recommended endpoint and copy-ready code.

Which endpoint should gemini-3.5-flash use?

Use https://api.tokenlab.sh/v1beta/models/gemini-3.5-flash:generateContent as the default for gemini-3.5-flash. If a provider-native format is supported, the API workbench shows that endpoint too.

Can I test gemini-3.5-flash before integrating it?

Yes. Try in Web Agent opens a ready test for gemini-3.5-flash and keeps your prompt after sign-in, so you don’t lose context.

Related Models