Gemini API Pricing and Cost per Workflow
By Paige Gilmore, Founder, NetLift · Published 2026-07-28 · Updated 2026-08-07
Prices checked 2026-08-02 against official vendor pricing pages.
Gemini API costs are structured per million tokens, with Gemini 3.1 Flash-Lite starting at $0.25 for input and Gemini 3.1 Pro Preview scaling from $2 to $18 depending on context length.
Gemini API pricing scales based on model choice and prompt volume. The Gemini 3.1 Flash-Lite model provides the lowest entry point with text and image input at $0.25 per 1M tokens. Organizations can optimize for different use cases by selecting between Standard, Priority, or Batch tiers depending on their requirements for latency and throughput.
For high-capacity needs, Gemini 3.1 Pro Preview introduces a tiered cost structure based on context length. Input and output rates increase significantly once a prompt exceeds 200,000 tokens, making context management a primary factor in total workflow expenditure.
Official per-token prices
| Model / rate | Price | Unit |
|---|---|---|
| Gemini 3.1 Flash-Lite — Standard Paid Tier - Input (Text/Image/Video) | $0.25 | per 1M input tokens |
| Gemini 3.1 Flash-Lite — Standard Paid - Input (Text/Image/Video) | $0.25 | per 1M input tokens |
| Gemini 3.1 Flash-Lite — Standard Paid Tier - Input (Audio) | $0.5 | per 1M input tokens |
| Gemini 3.1 Flash-Lite — Standard Paid - Input (Audio) | $0.5 | per 1M input tokens |
| Gemini 3.1 Flash-Lite — Standard Paid Tier - Output | $1.50 | per 1M output tokens |
| Gemini 3.1 Pro Preview — Standard Paid Tier - Input (prompts <= 200k tokens) | $2 | per 1M input tokens |
| Gemini 3.1 Pro Preview — Standard Paid - Input (<= 200k tokens) | $2 | per 1M input tokens |
| Gemini 3.1 Pro Preview — Standard Paid Tier - Input (prompts > 200k tokens) | $4 | per 1M input tokens |
| Gemini 3.1 Pro Preview — Standard Paid - Input (> 200k tokens) | $4 | per 1M input tokens |
| Gemini 3.1 Pro Preview — Standard Paid Tier - Output (prompts <= 200k tokens) | $12 | per 1M output tokens |
| Gemini 3.1 Pro Preview — Standard Paid - Output (<= 200k tokens) | $12 | per 1M output tokens |
| Gemini 3.1 Pro Preview — Standard Paid Tier - Output (prompts > 200k tokens) | $18 | per 1M output tokens |
| Gemini 3.1 Pro Preview — Standard Paid - Output (> 200k tokens) | $18 | per 1M output tokens |
| Gemini 3.5 Flash — Standard Paid Tier - Input | $1.50 | per 1M input tokens |
| Gemini 3.5 Flash — Standard Paid - Input | $1.50 | per 1M input tokens |
| Gemini 3.5 Flash — Standard Paid Tier - Output | $9 | per 1M output tokens |
| Gemini 3.5 Flash — Standard Paid - Output | $9 | per 1M output tokens |
| Gemini 3.5 Flash — Priority Paid Tier - Output | $16.20 | per 1M output tokens |
| Gemini 3.5 Flash-Lite — Standard Paid Tier - Input | $0.3 | per 1M input tokens |
| Gemini 3.5 Flash-Lite — Standard Paid - Input | $0.3 | per 1M input tokens |
| Gemini 3.5 Flash-Lite — Standard Paid Tier - Output | $2.50 | per 1M output tokens |
| Gemini 3.5 Flash-Lite — Standard Paid - Output | $2.50 | per 1M output tokens |
| Gemini 3.5 Live Translate — Paid Tier - Input | $3.50 | per 1M input tokens |
| Gemini 3.5 Live Translate — Paid Tier - Output | $21 | per 1M output tokens |
| Gemini 3.6 Flash — Standard Paid - Context Caching | $0.15 | per 1K tokens |
| Gemini 3.6 Flash — Batch/Flex Paid Tier - Input | $0.75 | per 1M input tokens |
| Gemini 3.6 Flash — Batch/Flex Paid - Input | $0.75 | per 1M input tokens |
| Gemini 3.6 Flash — Standard Paid Tier - Input | $1.50 | per 1M input tokens |
| Gemini 3.6 Flash — Standard Paid - Input | $1.50 | per 1M input tokens |
| Gemini 3.6 Flash — Priority Paid Tier - Input | $2.70 | per 1M input tokens |
| Gemini 3.6 Flash — Priority Paid - Input | $2.70 | per 1M input tokens |
| Gemini 3.6 Flash — Batch/Flex Paid Tier - Output | $3.75 | per 1M output tokens |
| Gemini 3.6 Flash — Batch/Flex Paid - Output | $3.75 | per 1M output tokens |
| Gemini 3.6 Flash — Standard Paid Tier - Output | $7.50 | per 1M output tokens |
| Gemini 3.6 Flash — Standard Paid - Output | $7.50 | per 1M output tokens |
| Gemini 3.6 Flash — Priority Paid Tier - Output | $13.50 | per 1M output tokens |
| Gemini 3.6 Flash — Priority Paid - Output | $13.50 | per 1M output tokens |
| Gemini Omni Flash Preview — Standard Paid Tier - Input | $1.50 | per 1M input tokens |
| Gemini Omni Flash Preview — Standard Paid - Input | $1.50 | per 1M input tokens |
| Gemini Omni Flash Preview — Standard Paid Tier - Output (Text) | $9 | per 1M output tokens |
How does context length impact Pro pricing?
Gemini 3.1 Pro Preview costs are tied directly to the size of the prompt. Standard inputs up to 200,000 tokens cost $2 per 1M, but this price doubles to $4 per 1M tokens for prompts exceeding that threshold. Output costs follow a similar logic, jumping from $12 to $18 per 1M tokens for larger context windows.
Which tier offers the lowest unit cost for high-volume tasks?
Gemini 3.6 Flash offers a Batch/Flex tier designed for high-volume efficiency, pricing input at $0.75 and output at $3.75 per 1M tokens. This represents a substantial discount compared to its Priority tier, which charges $2.70 for input and $13.50 for output for the same model. Additionally, Gemini 3.6 Flash supports context caching at a rate of $0.15 per 1K tokens to further manage recurring prompt costs.
Determining the ROI of your Gemini implementation requires moving beyond simple token tallies to measure actual workflow efficiency. By comparing the cost of model inference against the time saved in manual engineering or administrative tasks, you can establish a clear payback period. NetLift enables teams to track these unit economics, ensuring that expensive Pro calls deliver a net return compared to lower-cost Flash alternatives.
Sources
All pricing on this page comes from official vendor pages:
- https://ai.google.dev/gemini-api/docs/pricing — retrieved 2026-07-27
Frequently asked questions
How do premium AI tools like Gemini 3.1 Pro compare in pricing?
Gemini 3.1 Pro Preview costs $2 per 1M input tokens for prompts up to 200k tokens. For larger prompts over 200k tokens, the price increases to $4 per 1M input tokens.
How can marketers win in the Gemini era by optimizing costs?
Marketers can minimize overhead by using Gemini 3.5 Flash-Lite for text tasks, which costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, rather than using more expensive Pro models.
What are the costs for Gemini 3.5 Live Translate?
The paid tier for Gemini 3.5 Live Translate is priced at $3.50 per 1M input tokens and $21 per 1M output tokens.
What is the cheapest option for video or image input?
Gemini 3.1 Flash-Lite offers the lowest rate for image and video input at $0.25 per 1M tokens under the Standard Paid Tier.
About the author
Paige Gilmore is the founder of NetLift, the AI Value Management platform that helps organisations measure the cost, savings and return of AI adoption. Paige Gilmore on LinkedIn