Google Gemini API Pricing
Compare Google Gemini API pricing across 18 models, from $0.075/1M to $2.00/1M input tokens. See live costs for Gemini 3.1 Pro, Gemini 2.5 Pro, Gemini 2.5 Flash, Gemini 2.0 Flash, and the latest Gemini Flash-Lite options.
Google pricing deep dives
Need the exact Google Gemini API pricing breakdown instead of the full provider table? Start with the dedicated Gemini guide, then use the Gemma analysis for Google's open-model pricing.
Google Gemini Recommendation
Compare Google Gemini API pricing across 18 models, from $0.075/1M to $2.00/1M input tokens. See live costs for Gemini 3.1 Pro, Gemini 2.5 Pro, Gemini 2.5 Flash, Gemini 2.0 Flash, and the latest Gemini Flash-Lite options.
Compare Google Gemini vs other providers
Google Models & Pricing
| Model | Input Price | Output Price | Context Window | Max Output | Category | Capabilities |
|---|---|---|---|---|---|---|
Gemini 3.8 Live Extended Thinking Google's high-reasoning live audio model for complex, multi-step voice agents with background reasoning and asynchronous tool use. Released: 9/15/2026 | $0.75 per 1M tokens | $4.50 per 1M tokens | 131,072 tokens | 65,536 tokens | reasoning | textvisionaudiovideoreasoningstreaming |
Gemini 3.8 Live Google's low-latency audio-to-audio model for real-time voice agents, live dialogue, visual grounding, and asynchronous tool use. Released: 9/15/2026 | $0.75 per 1M tokens | $4.50 per 1M tokens | 131,072 tokens | 65,536 tokens | balanced | textvisionaudiovideoreasoningstreaming |
Gemini 3.8 Flash Google's most intelligent Flash model for long-horizon software engineering, autonomous agents, and complex enterprise workflows. Released: 9/2/2026 | $0.75 per 1M tokens | $3.75 per 1M tokens | 1,048,576 tokens | 65,536 tokens | flagship | textvisionaudiovideocodereasoning |
Gemini Omni 1.1 Flash Google's generally available conversational video generation and editing model with native audio, interpolation, extension, and resolution control. Released: 8/27/2026 | $1.50 per 1M tokens | $17.50 per 1M tokens | 1,048,576 tokens | 57,920 tokens | flagship | textvisionaudiovideo |
Gemini 3.7 Flash Google's most capable Flash workhorse for agentic coding, multimodal reasoning, and reliable multi-step execution. Released: 8/13/2026 | $0.75 per 1M tokens | $3.75 per 1M tokens | 1,000,000 tokens | 65,536 tokens | flagship | textvisionaudiovideocodereasoning |
Gemini 3.6 Flash Google's latest balanced Flash-tier model for agentic coding, multimodal work, and faster production loops. Released: 7/21/2026 | $1.50 per 1M tokens | $7.50 per 1M tokens | 1,000,000 tokens | 65,536 tokens | balanced | textvisionaudiovideocodereasoning |
Gemini 3.5 Flash-Lite Google's low-latency, cost-efficient Flash-Lite model for high-throughput agents, extraction, and parsing. Released: 7/21/2026 | $0.30 per 1M tokens | $2.50 per 1M tokens | 1,000,000 tokens | 65,536 tokens | efficient | textvisionaudiovideocodereasoning |
Gemini 3.5 Flash Frontier-level Gemini model optimized for agents, coding, and long-horizon workflows. Released: 5/19/2026 | $1.50 per 1M tokens | $9.00 per 1M tokens | 1,000,000 tokens | 65,536 tokens | flagship | textvisionaudiovideocodereasoning |
Gemini Embedding 2 First natively multimodal embedding model supporting text, images, video, audio, and documents Released: 3/10/2026 | $0.20 per 1M tokens | $0.20 per 1M tokens | 8,192 tokens | 3,072 tokens | embedding | textvisionaudiovideoembeddings |
Gemini 3.1 Flash-Lite Preview Most cost-effective model in the Gemini 3.1 family, optimized for high-volume workloads Released: 3/3/2026 | $0.25 per 1M tokens | $1.50 per 1M tokens | 1,000,000 tokens | 8,192 tokens | efficient | textvisioncode |
Gemini 3.1 Pro Latest Gemini 3 series with improved intelligence, multimodal understanding, and agentic capabilities Released: 2/19/2026 | $2.00 per 1M tokens | $12.00 per 1M tokens | 1,000,000 tokens | 65,536 tokens | flagship | textvisionaudiovideocodereasoning |
$0.50 per 1M tokens | $3.00 per 1M tokens | 1,000,000 tokens | 65,536 tokens | efficient | textvisionaudiocode | |
Gemini 3 Pro Latest frontier Gemini model with exceptional multimodal capabilities Released: 11/18/2025 | $2.00 per 1M tokens | $12.00 per 1M tokens | 2,000,000 tokens | 131,072 tokens | flagship | textvisionaudiovideocode |
Gemini 2.5 Flash-Lite Smallest and most cost-effective Gemini model for at-scale usage Released: 6/17/2025 | $0.10 per 1M tokens | $0.40 per 1M tokens | 1,000,000 tokens | 32,768 tokens | efficient | textvisionaudio |
$0.30 per 1M tokens | $2.50 per 1M tokens | 1,000,000 tokens | 32,768 tokens | efficient | textvisionaudiocode | |
$1.25 per 1M tokens | $10.00 per 1M tokens | 2,000,000 tokens | 131,072 tokens | flagship | textvisionaudiovideocode | |
$0.075 per 1M tokens | $0.30 per 1M tokens | 1,000,000 tokens | 32,768 tokens | efficient | textvisionaudio | |
$0.10 per 1M tokens | $0.40 per 1M tokens | 1,000,000 tokens | 32,768 tokens | efficient | textvisionaudiocode |
Calculate Google Costs
Use our calculator to estimate costs for any Google model based on your usage patterns.
Compare Google Models
Gemini 3.8 Live Extended Thinking vs Gemini 3.8 Live
Gemini 3.8 Live vs Gemini 3.6 Flash
Gemini 3.8 Flash vs Gemini Omni 1.1 Flash
Gemini Omni 1.1 Flash vs Gemini 3.7 Flash
Gemini 3.7 Flash vs Gemini 3.5 Flash
Gemini 3.6 Flash vs Gemini 3.5 Flash-Lite
Google Pricing FAQ
How much does Google Gemini API cost?
Google Gemini pricing on this page ranges from $0.075 to $2.00 per million input tokens across 18 models. The cheapest option is Gemini 2.0 Flash-Lite at $0.075/1M input and $0.30/1M output tokens.
What is the cheapest Google Gemini model?
Gemini 2.0 Flash-Lite is the cheapest Google Gemini model here at $0.075 per million input tokens and $0.30 per million output tokens. It supports a 1,000,000-token context window.
Which Gemini model should I use?
Gemini 2.0 Flash-Lite is the best low-cost Google Gemini option on this page. Gemini 3.1 Pro ($2.00/1M input) is the premium pick when quality matters more than token spend. Gemini Flash models are usually the sweet spot for high-volume workloads. For heavier reasoning tasks, start with the higher-end Gemini Pro models.
What Gemini models are available?
This page compares 18 Google Gemini models: Gemini 3.8 Live, Gemini 3.8 Live Extended Thinking, Gemini Omni 1.1 Flash, Gemini 3.8 Flash, Gemini 3.7 Flash, Gemini 3.5 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.1 Flash-Lite Preview, Gemini 3.1 Pro, Gemini 3 Pro, Gemini 3 Flash, Gemini 2.5 Pro, Gemini 2.5 Flash, Gemini 2.0 Flash, Gemini 2.5 Flash-Lite, Gemini 2.0 Flash-Lite, Gemini Embedding 2. These span balanced, reasoning, flagship, efficient, embedding categories with context windows up to 2000K tokens.
How does Google Gemini pricing compare to OpenAI or Anthropic?
Google Gemini pricing starts at $0.075/1M input tokens with Gemini 2.0 Flash-Lite. Use our comparison pages to stack Gemini against OpenAI, Anthropic, and Mistral on token cost, context window, and model tier.
Related Articles
AX Agent Orchestrator: 6 Production Workflows Founders and Operators Can Build Now
How to use Google's AX agent orchestrator for traceable research, support, coding, data enrichment, and back-office w...
9/21/2026
What Gemini 3.8 Live Makes Possible: 6 Real-Time Multimodal Workflows to Build Now
Gemini 3.8 Live pricing and workflow guide for voice-and-screen AI in support, sales, incidents, meetings, and QA wit...
9/16/2026
Google's /goto Update Broke Fragile AI Scrapers: How to Rebuild Reliable Research Agents
Google /goto links are breaking brittle scrapers. Rebuild AI research agents with URL normalization, evidence capture...
9/12/2026
Gemini Omni Flash and Nano Banana 2 Lite: 6 Creative Workflows Teams Can Ship Now
Google's Gemini Omni Flash and Nano Banana 2 Lite open practical image-to-video workflows for creative teams, agencie...
7/27/2026