Gemini API Pricing: Flash & Pro Cost Analysis

Gemini API Pricing and Cost per Workflow

By Paige Gilmore, Founder, NetLift · Published 2026-07-28 · Updated 2026-08-07

Prices checked 2026-08-02 against official vendor pricing pages.

Gemini API costs are structured per million tokens, with Gemini 3.1 Flash-Lite starting at $0.25 for input and Gemini 3.1 Pro Preview scaling from $2 to $18 depending on context length.

Gemini API pricing scales based on model choice and prompt volume. The Gemini 3.1 Flash-Lite model provides the lowest entry point with text and image input at $0.25 per 1M tokens. Organizations can optimize for different use cases by selecting between Standard, Priority, or Batch tiers depending on their requirements for latency and throughput.

For high-capacity needs, Gemini 3.1 Pro Preview introduces a tiered cost structure based on context length. Input and output rates increase significantly once a prompt exceeds 200,000 tokens, making context management a primary factor in total workflow expenditure.

Official per-token prices

Model / rate Price Unit
Gemini 3.1 Flash-Lite — Standard Paid Tier - Input (Text/Image/Video) $0.25 per 1M input tokens
Gemini 3.1 Flash-Lite — Standard Paid - Input (Text/Image/Video) $0.25 per 1M input tokens
Gemini 3.1 Flash-Lite — Standard Paid Tier - Input (Audio) $0.5 per 1M input tokens
Gemini 3.1 Flash-Lite — Standard Paid - Input (Audio) $0.5 per 1M input tokens
Gemini 3.1 Flash-Lite — Standard Paid Tier - Output $1.50 per 1M output tokens
Gemini 3.1 Pro Preview — Standard Paid Tier - Input (prompts <= 200k tokens) $2 per 1M input tokens
Gemini 3.1 Pro Preview — Standard Paid - Input (<= 200k tokens) $2 per 1M input tokens
Gemini 3.1 Pro Preview — Standard Paid Tier - Input (prompts > 200k tokens) $4 per 1M input tokens
Gemini 3.1 Pro Preview — Standard Paid - Input (> 200k tokens) $4 per 1M input tokens
Gemini 3.1 Pro Preview — Standard Paid Tier - Output (prompts <= 200k tokens) $12 per 1M output tokens
Gemini 3.1 Pro Preview — Standard Paid - Output (<= 200k tokens) $12 per 1M output tokens
Gemini 3.1 Pro Preview — Standard Paid Tier - Output (prompts > 200k tokens) $18 per 1M output tokens
Gemini 3.1 Pro Preview — Standard Paid - Output (> 200k tokens) $18 per 1M output tokens
Gemini 3.5 Flash — Standard Paid Tier - Input $1.50 per 1M input tokens
Gemini 3.5 Flash — Standard Paid - Input $1.50 per 1M input tokens
Gemini 3.5 Flash — Standard Paid Tier - Output $9 per 1M output tokens
Gemini 3.5 Flash — Standard Paid - Output $9 per 1M output tokens
Gemini 3.5 Flash — Priority Paid Tier - Output $16.20 per 1M output tokens
Gemini 3.5 Flash-Lite — Standard Paid Tier - Input $0.3 per 1M input tokens
Gemini 3.5 Flash-Lite — Standard Paid - Input $0.3 per 1M input tokens
Gemini 3.5 Flash-Lite — Standard Paid Tier - Output $2.50 per 1M output tokens
Gemini 3.5 Flash-Lite — Standard Paid - Output $2.50 per 1M output tokens
Gemini 3.5 Live Translate — Paid Tier - Input $3.50 per 1M input tokens
Gemini 3.5 Live Translate — Paid Tier - Output $21 per 1M output tokens
Gemini 3.6 Flash — Standard Paid - Context Caching $0.15 per 1K tokens
Gemini 3.6 Flash — Batch/Flex Paid Tier - Input $0.75 per 1M input tokens
Gemini 3.6 Flash — Batch/Flex Paid - Input $0.75 per 1M input tokens
Gemini 3.6 Flash — Standard Paid Tier - Input $1.50 per 1M input tokens
Gemini 3.6 Flash — Standard Paid - Input $1.50 per 1M input tokens
Gemini 3.6 Flash — Priority Paid Tier - Input $2.70 per 1M input tokens
Gemini 3.6 Flash — Priority Paid - Input $2.70 per 1M input tokens
Gemini 3.6 Flash — Batch/Flex Paid Tier - Output $3.75 per 1M output tokens
Gemini 3.6 Flash — Batch/Flex Paid - Output $3.75 per 1M output tokens
Gemini 3.6 Flash — Standard Paid Tier - Output $7.50 per 1M output tokens
Gemini 3.6 Flash — Standard Paid - Output $7.50 per 1M output tokens
Gemini 3.6 Flash — Priority Paid Tier - Output $13.50 per 1M output tokens
Gemini 3.6 Flash — Priority Paid - Output $13.50 per 1M output tokens
Gemini Omni Flash Preview — Standard Paid Tier - Input $1.50 per 1M input tokens
Gemini Omni Flash Preview — Standard Paid - Input $1.50 per 1M input tokens
Gemini Omni Flash Preview — Standard Paid Tier - Output (Text) $9 per 1M output tokens

How does context length impact Pro pricing?

Gemini 3.1 Pro Preview costs are tied directly to the size of the prompt. Standard inputs up to 200,000 tokens cost $2 per 1M, but this price doubles to $4 per 1M tokens for prompts exceeding that threshold. Output costs follow a similar logic, jumping from $12 to $18 per 1M tokens for larger context windows.

Which tier offers the lowest unit cost for high-volume tasks?

Gemini 3.6 Flash offers a Batch/Flex tier designed for high-volume efficiency, pricing input at $0.75 and output at $3.75 per 1M tokens. This represents a substantial discount compared to its Priority tier, which charges $2.70 for input and $13.50 for output for the same model. Additionally, Gemini 3.6 Flash supports context caching at a rate of $0.15 per 1K tokens to further manage recurring prompt costs.

Determining the ROI of your Gemini implementation requires moving beyond simple token tallies to measure actual workflow efficiency. By comparing the cost of model inference against the time saved in manual engineering or administrative tasks, you can establish a clear payback period. NetLift enables teams to track these unit economics, ensuring that expensive Pro calls deliver a net return compared to lower-cost Flash alternatives.

Sources

All pricing on this page comes from official vendor pages:

Frequently asked questions

How do premium AI tools like Gemini 3.1 Pro compare in pricing?

Gemini 3.1 Pro Preview costs $2 per 1M input tokens for prompts up to 200k tokens. For larger prompts over 200k tokens, the price increases to $4 per 1M input tokens.

How can marketers win in the Gemini era by optimizing costs?

Marketers can minimize overhead by using Gemini 3.5 Flash-Lite for text tasks, which costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, rather than using more expensive Pro models.

What are the costs for Gemini 3.5 Live Translate?

The paid tier for Gemini 3.5 Live Translate is priced at $3.50 per 1M input tokens and $21 per 1M output tokens.

What is the cheapest option for video or image input?

Gemini 3.1 Flash-Lite offers the lowest rate for image and video input at $0.25 per 1M tokens under the Standard Paid Tier.

About the author

Paige Gilmore is the founder of NetLift, the AI Value Management platform that helps organisations measure the cost, savings and return of AI adoption.

Related resources

Start a free NetLift trial · See a sample report · All resources