Gemini API Pricing: Cost Analysis for 3.5 and 3.8 Models
AI Costs & Pricing

Gemini API Pricing and Cost per Workflow

By Paige Gilmore, Founder, NetLift· Published July 28, 2026· Updated September 25, 2026
Prices checked September 25, 2026 against official vendor pages

Gemini API pricing is segmented by model version and processing tier, with Gemini 3.8 Flash offering lower entry costs than 3.5 Flash for standard text workflows. Costs scale significantly for live audio features, while context caching and batch tiers provide the most aggressive discounts for high-volume users.

14-day free trial. No credit card required.

Gemini API costs depend on the model version (3.5 vs 3.8) and the processing urgency required for the workflow. For engineering teams, the most significant savings are found in the Gemini 3.8 Flash Batch tier, which reduces input costs compared to standard processing.

Multimodal workflows require more granular budgeting, as Gemini 3.8 Live distinguishes between text, image, and audio inputs. While text remains the most affordable medium, audio inputs and outputs carry a premium, making it essential to audit media-heavy agent spend against traditional text-based automation.

Official per-token prices

Model / rate Price Unit
Gemini 3.5 Flash — Batch Paid Tier Input $0.75 per 1M input tokens
Gemini 3.5 Flash — Standard Paid Tier - Input $1.50 per 1M input tokens
Gemini 3.5 Flash — Standard Paid Output $9 per 1M output tokens
Gemini 3.5 Live Translate — Standard Paid Tier - Input (audio) $3.50 per 1M input tokens
Gemini 3.5 Live Translate — Standard Paid Tier - Output (audio) $21 per 1M output tokens
Gemini 3.8 Flash — Standard Paid Tier Context Caching $0.07 per 1M input tokens
Gemini 3.8 Flash — Batch Paid Tier Input $0.38 per 1M input tokens
Gemini 3.8 Flash — Standard Paid Tier Input $0.75 per 1M input tokens
Gemini 3.8 Flash — Priority Paid Tier Input $1.35 per 1M input tokens
Gemini 3.8 Flash — Standard Paid Tier Output $3.75 per 1M output tokens
Gemini 3.8 Live — Standard Paid Tier Input (Text) $0.75 per 1M input tokens
Gemini 3.8 Live — Standard Paid Tier Input (Image/Video) $1 per 1M input tokens
Gemini 3.8 Live — Standard Paid Tier Input (Audio) $3 per 1M input tokens
Gemini 3.8 Live — Standard Paid Tier Output (Text) $4.50 per 1M output tokens
Gemini 3.8 Live — Standard Paid Tier Output (Audio) $12 per 1M output tokens

14-day free trial. No credit card required.

Which tier offers the best unit economics for high-volume workflows?

For background processing and non-real-time data extraction, the Batch Paid Tier provides the most efficient pricing. Using Gemini 3.8 Flash in Batch mode reduces input costs compared to the Standard tier. Additionally, developers can leverage Context Caching for Gemini 3.8 Flash to handle large datasets at a fraction of the cost of fresh input tokens, provided the data remains in the cache.

How does model choice affect live and audio agent budgets?

Building voice or live translation agents requires accounting for higher output premiums. Gemini 3.5 Live Translate and Gemini 3.8 Live charge significantly more for audio output than for text. Organizations moving from text-only bots to voice-enabled agents should expect a substantial increase in per-token costs, particularly on the output side where audio generation is priced higher than standard text generation.

To determine if Gemini adoption is generating a positive return, teams must look past token rates to measure the net return on AI. This involves comparing the cost of API calls and engineering resources against the quantifiable time saved by the end user. NetLift enables organizations to map these Gemini 3.8 Flash costs to specific business outcomes, ensuring that the speed of the model translates into a measurable reduction in operational overhead.

14-day free trial. No credit card required.

Sources

All pricing on this page comes from official vendor pages:

Frequently Asked Questions

Ready to see your AI return?

14-day free trial. No credit card required.

About the author

Paige Gilmore · Founder, NetLift

Paige Gilmore is the founder of NetLift, the AI Value Management platform that helps organisations measure the cost, savings and return of AI adoption.

Paige Gilmore on LinkedIn

Keep reading