Skip to main content
⚡ EfficientBudget-Friendly

Gemini 3.5 Flash-Lite

by Google

Google's low-latency, cost-efficient Flash-Lite model for high-throughput agents, extraction, and parsing.

Try Google

Input Price

$0.3

per 1M tokens

Output Price

$2.5

per 1M tokens

Context Window

1,000K

tokens (65.536K max output)

Specifications

Release Date2026-07-21
Category⚡ Efficient
Context Window1,000,000 tokens
Max Output65,536 tokens
ProviderGoogle
Input Cost$0.3 / 1M tokens
Output Cost$2.5 / 1M tokens
Price TierBudget-Friendly

Capabilities

textvisionaudiovideocodereasoning

Related Use Cases

Monthly Cost Estimates

Usage LevelDaily TokensDailyMonthlyYearly
Light~50 requests/day100K in / 20K out$0.08$2.40$29
Medium~200 requests/day500K in / 100K out$0.40$12.00$146
Heavy~1K requests/day2,000K in / 500K out$1.85$55.50$675
Enterprise~5K requests/day10,000K in / 2,000K out$8.00$240.00$2920

Cost Calculator

Alternatives to Gemini 3.5 Flash-Lite

Frequently Asked Questions

How much does Gemini 3.5 Flash-Lite API cost per million tokens?
Gemini 3.5 Flash-Lite costs $0.3 per million input tokens and $2.5 per million output tokens as of 2026. These are the standard API rates from Google.
What is the Gemini 3.5 Flash-Lite context window?
Gemini 3.5 Flash-Lite supports a 1M context window (1,000,000 tokens), which means you can process up to 1M tokens in a single API call.
How much does Gemini 3.5 Flash-Lite cost per month?
At medium usage (~200 requests/day with 500K input and 100K output tokens/day), Gemini 3.5 Flash-Lite costs approximately $12.00/month. Light usage runs about $2.40/month, and heavy usage (~1K requests/day) around $55.50/month.
Is there a cheaper alternative to Gemini 3.5 Flash-Lite?
Yes — Devstral 2 by Mistral AI is a cheaper option at $0.4/M input tokens vs $0.3/M for Gemini 3.5 Flash-Lite. Other budget alternatives include models in the efficient tier.
Is Gemini 3.5 Flash-Lite good for ai agent / agentic workflows?
Gemini 3.5 Flash-Lite is a efficient model with support for text, vision, audio. For ai agent / agentic workflows, it offers 1M context and costs $0.3/M input tokens — a budget-friendly choice for this use case.

Gemini 3.5 Flash-Lite Comparisons

Ready to use Gemini 3.5 Flash-Lite?

Get started with Google's API — free tier available for most models.

Try Google API →