Choose Auto, TokenLab Verified, or Official for each request, with prices shown up front.See what's new

Back to models

Gemini 3.5 Flash vs GPT-5.4 mini

Gemini 3.5 Flash and GPT-5.4 mini: current price, context, max output, and supported operations.

Pricing comparison

Gemini 3.5 FlashGPT-5.4 mini
Model makerGoogleOpenAI
Delivery availabilityAvailableAvailable
Context window1M400K
Max output64K128K
Official price
Input$1.50per 1M tokensOutput$9.00per 1M tokens
Input$0.75per 1M tokensOutput$4.50per 1M tokens
Verified price
Input$0.75per 1M tokensOutput$4.50per 1M tokens
Input$0.525per 1M tokensOutput$3.15per 1M tokens
Model performance
30-day success rate
94.8%
30-day success rate
98.8%
7-day median latency
5.1 sn=968
Capabilities
JSON modePrompt CacheReasoningTool useVision
Prompt CacheReasoningTool useVision

Summary

  • Gemini 3.5 Flash: Input $0.75 / Output $4.50 per 1M tokens; GPT-5.4 mini: Input $0.525 / Output $3.15 per 1M tokens. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
  • Gemini 3.5 Flash has a 2.6x larger context window
  • GPT-5.4 mini supports 2.0x more output tokens
Gemini 3.5 Flash
Gemini 3
View details
GPT-5.4 mini
GPT 5.4
View details

FAQ

Which is cheaper, Gemini 3.5 Flash or GPT-5.4 mini?

Gemini 3.5 Flash: Input $0.75 / Output $4.50 per 1M tokens; GPT-5.4 mini: Input $0.525 / Output $3.15 per 1M tokens. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What are the key differences between Gemini 3.5 Flash and GPT-5.4 mini?

Gemini 3.5 Flash: JSON mode