Google

Gemma 4 26b A4b API Pricing

Gemma 4 26b A4b costs $0.13 per 1M input tokens and $0.4 per 1M output tokens on Google's API. Cached input is billed at $0.085 per 1M tokens. Prices are provider list prices in USD, refreshed as sources update; see the Gemma 4 26b A4b model profile for capability scores.

Input / 1M
$0.13
Output / 1M
$0.4
Effective / 1M I/O
$0.4462
Cost Rank
#12

Price List

RatePrice (USD)Notes
Input$0.13per 1M tokens
Output$0.4per 1M tokens
Cached input (read)$0.085per 1M tokens
Blended input + output$0.531M in + 1M out at list price
Context window256K tokens

What a Workload Costs

Computed straight from the list prices above — token counts are illustrative workload sizes, not measurements of Gemma 4 26b A4b.

WorkloadCostTokens
Short chat turn$0.000331K in / 500 out
Summarize a long document$0.014100K in / 2K out
Agentic coding session$0.105500K in / 100K out
1M input + 1M output tokens$0.5301M in / 1M out

Effective Cost

In AI IQ's scoring, Gemma 4 26b A4b's list price is adjusted by a token-usage multiplier of 0.842 (measured from real benchmark runs), giving an effective cost of $0.4462 per 1M input + output tokens. That places it #12 of 117 models on the effective-cost ranking (lower is cheaper). Some models spend far more tokens than others on the same task, so effective cost compares what a unit of work really costs. Method details are on the methodology page; see all models on the cost charts and the falling cost of intelligence over time.

Cheaper Alternatives at Similar Capability

Models with a lower effective cost that score within a few IQ points of Gemma 4 26b A4b (IQ 96), or better.

ModelProviderIQEffective Cost / 1M
gemma-4-31bGoogle101$0.3359
gpt-oss-20bOpenAI100$0.1937
gemma-4-12bGoogle98$0.2859
nemotron-3-nanoUS (Other)96$0.2732
gpt-5-nanoOpenAI96$0.303
Nemotron 3.5 LightningUS (Other)95$0.3284

Compare