Google

Gemini 3.5 Flash Lite API Pricing

Gemini 3.5 Flash Lite costs $0.3 per 1M input tokens and $2.50 per 1M output tokens on Google's API. Prices are provider list prices in USD, refreshed as sources update; see the Gemini 3.5 Flash Lite model profile for capability scores.

Input / 1M
$0.3
Output / 1M
$2.50
Effective / 1M I/O
$3.60
Cost Rank
#54

Price List

RatePrice (USD)Notes
Input$0.3per 1M tokens
Output$2.50per 1M tokens
Blended input + output$2.801M in + 1M out at list price
Context window1M tokens

What a Workload Costs

Computed straight from the list prices above — token counts are illustrative workload sizes, not measurements of Gemini 3.5 Flash Lite.

WorkloadCostTokens
Short chat turn$0.001551K in / 500 out
Summarize a long document$0.035100K in / 2K out
Agentic coding session$0.400500K in / 100K out
1M input + 1M output tokens$2.801M in / 1M out

Effective Cost

In AI IQ's scoring, Gemini 3.5 Flash Lite's list price is adjusted by a token-usage multiplier of 1.285 (measured from real benchmark runs), giving an effective cost of $3.60 per 1M input + output tokens. That places it #54 of 117 models on the effective-cost ranking (lower is cheaper). Some models spend far more tokens than others on the same task, so effective cost compares what a unit of work really costs. Method details are on the methodology page; see all models on the cost charts and the falling cost of intelligence over time.

Cheaper Alternatives at Similar Capability

Models with a lower effective cost that score within a few IQ points of Gemini 3.5 Flash Lite (IQ 108), or better.

ModelProviderIQEffective Cost / 1M
mimo-v2.5-proXiaomi116$1.80
gemini-3-flashGoogle116$2.74
qwen3.7-plusAlibaba115$1.67
grok-4.3xAI115$2.60
glm-5.1Z.ai114$3.14
deepseek-v4-flashDeepSeek113$1.39

Compare