Choose Auto, TokenLab Verified, or Official for each request, with prices shown up front.See what's new

Back to models

Gemini 3.5 Flash-Lite vs GPT-5.4 mini

Gemini 3.5 Flash-Lite and GPT-5.4 mini: current price, context, max output, and supported operations.

Pricing comparison

Gemini 3.5 Flash-LiteGPT-5.4 mini
Model makerGoogleOpenAI
Delivery availabilityAvailableAvailable
Context window1M400K
Max output64K128K
Official price
Input$0.30per 1M tokensOutput$2.50per 1M tokens
Input$0.75per 1M tokensOutput$4.50per 1M tokens
TokenLab price
Input$0.15per 1M tokensOutput$1.25per 1M tokens
Input$0.525per 1M tokensOutput$3.15per 1M tokens
Model performance
30-day success rate
99.7%
7-day median latency
1.9 sn=346
30-day success rate
98.3%
Capabilities
JSON modePrompt CacheTool useVision
Prompt CacheReasoningTool useVision

Summary

  • Gemini 3.5 Flash-Lite: Input $0.15 / Output $1.25 per 1M tokens; GPT-5.4 mini: Input $0.525 / Output $3.15 per 1M tokens. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
  • Gemini 3.5 Flash-Lite has a 2.6x larger context window
  • GPT-5.4 mini supports 2.0x more output tokens
Gemini 3.5 Flash-Lite
Gemini 3
View details
GPT-5.4 mini
GPT 5.4
View details

FAQ

Which is cheaper, Gemini 3.5 Flash-Lite or GPT-5.4 mini?

Gemini 3.5 Flash-Lite: Input $0.15 / Output $1.25 per 1M tokens; GPT-5.4 mini: Input $0.525 / Output $3.15 per 1M tokens. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What are the key differences between Gemini 3.5 Flash-Lite and GPT-5.4 mini?

Gemini 3.5 Flash-Lite: JSON mode. GPT-5.4 mini: Reasoning