Google Gemini API Pricing 2026
Complete pricing guide with plans, hidden costs, and cost analysis
Google Gemini API pricing ranges from $0 to $18 per million tokens.
Are you Google Gemini API? Claim this profile
Is Google Gemini API a fit?
Derived from this product's own pricing and contract record — not a review.
Best for
- Developers and small projects getting started with the Gemini API
- High-volume, cost-sensitive production workloads
- Production apps balancing cost, speed, and capability
- Complex reasoning, long-context, and multimodal tasks
Not a fit if
- The sticker price is your whole budget — Output Reliability and Accuracy Costs (and 1 more).
All Google Gemini API Plans & Pricing
| Plan | Monthly | Annual | Best For |
|---|---|---|---|
| Free rate_limit: Free tier rate limits apply | Free | Free | Developers and small projects getting started with the Gemini API |
| Verified pricing · last checked August 2026 · 2 sources
Get this price at Google Gemini API →
| |||
| What's included at Free Best for: Developers and small projects getting started with the Gemini API
Limits
| |||
| Flash-Lite (Paid) | Contact Sales | Contact Sales | High-volume, cost-sensitive production workloads |
| Verified pricing · last checked August 2026 · 2 sources
Get this price at Google Gemini API →
| |||
| What's included at Flash-Lite (Paid) Best for: High-volume, cost-sensitive production workloads
| |||
| Flash (Paid) | Contact Sales | Contact Sales | Production apps balancing cost, speed, and capability |
| Verified pricing · last checked August 2026 · 2 sources
Get this price at Google Gemini API →
| |||
| What's included at Flash (Paid) Best for: Production apps balancing cost, speed, and capability
| |||
| Pro (Paid) | Contact Sales | Contact Sales | Complex reasoning, long-context, and multimodal tasks |
| Verified pricing · last checked August 2026 · 2 sources
Get this price at Google Gemini API →
| |||
| What's included at Pro (Paid) Best for: Complex reasoning, long-context, and multimodal tasks
| |||
View all features by plan (compare side-by-side)
Free
- Limited access to certain models
- Free input & output tokens
- Google AI Studio access
- Content used to improve Google products
Flash-Lite (Paid)
- Gemini 2.5 Flash-Lite: $0.10 input / $0.40 output per 1M tokens
- Gemini 3.1 Flash-Lite: $0.25 input / $1.50 output per 1M tokens
- Gemini 3.5 Flash-Lite: $0.30 input / $2.50 output per 1M tokens
- Batch and Flex pricing available at reduced rates
- Context caching available where supported
Flash (Paid)
- Gemini 2.5 Flash: $0.30 input / $2.50 output per 1M tokens
- Gemini 3 Flash Preview: $0.50 input / $3.00 output per 1M tokens
- Gemini 3.5 Flash: $1.50 input / $9.00 output per 1M tokens
- Gemini 3.6 Flash: $1.50 input / $7.50 output per 1M tokens
- Batch, Flex, and Priority pricing available for supported models
Pro (Paid)
- Gemini 2.5 Pro: $1.25 input / $10.00 output per 1M tokens for prompts <=200K tokens
- Gemini 2.5 Pro: $2.50 input / $15.00 output per 1M tokens for prompts >200K tokens
- Gemini 3.1 Pro Preview: $2.00 input / $12.00 output per 1M tokens for prompts <=200K tokens
- Gemini 3.1 Pro Preview: $4.00 input / $18.00 output per 1M tokens for prompts >200K tokens
- Context caching available
Track Google Gemini API pricing
Get an email when Google Gemini API's pricing changes — plus the weekly SaaS Price Watch: verified price changes and deals across 3,000+ products. One-click unsubscribe.
You're on the list — first digest lands Tuesday.
Google Gemini API costs Free to $18 per million tokens as of August 2026, with 4 plans available including a free tier. Plan: Free (free). Custom pricing is available on request. Pricing depends on your chosen tier, contract length, and negotiated discounts.
Use the interactive pricing calculator to estimate your exact cost based on team size and requirements.
- Free tier: Yes
Google Gemini API offers 4 pricing tiers: Free, Flash-Lite (Paid), Flash (Paid), Pro (Paid). The Flash-Lite (Paid) plan is high-volume, cost-sensitive production workloads.
Compared to other llm api providers software, Google Gemini API is positioned at the budget-friendly price point.
- 4 documented hidden costs beyond list price
How much does Google Gemini API cost?
Google Gemini API Pricing Overview
Google Gemini API has 4 pricing plans, including a free tier. Paid plans range from $0 to $18/per million tokens. The Free plan is free and is best for developers and small projects getting started with the gemini api. The Flash-Lite (Paid) plan requires contacting sales for a custom quote and is designed for high-volume, cost-sensitive production workloads. The Flash (Paid) plan requires contacting sales for a custom quote and is designed for production apps balancing cost, speed, and capability. The Pro (Paid) plan requires contacting sales for a custom quote and is designed for complex reasoning, long-context, and multimodal tasks.
There are at least 4 documented hidden costs beyond Google Gemini API's list price, including implementation, training, and add-on fees.
This pricing was last verified in August 8, 2026 from 2 independent sources.
Google Gemini API pricing starts at $0 on the Free tier, which provides rate-limited access to Gemini models via AI Studio for prototyping. For production workloads, the Flash-Lite (Paid), Flash (Paid), and Pro (Paid) tiers are all billed on a per-token usage basis with no monthly subscription or minimum commitment. According to Artificial Analysis data from April 2026, the provider median across 51 tracked models sits at $0.56 per 1M input tokens and $2.20 per 1M output tokens, with model-level rates ranging from $0 (open Gemma models) up to $10.00 per 1M output tokens for Gemini 2.5 Pro.
How Google Gemini API Pricing Compares
Compare Google Gemini API pricing against top alternatives in LLM API Providers.
Usage-Based Rates
Per-unit pricing for Google Gemini API API usage.
Flash-Lite (Paid)
| Model | Input | Output | Cached | Per |
|---|---|---|---|---|
| gemini-2-5-flash-lite | $0.100 | $0.400 | $0.010 | 1M tokens |
| gemini-3-1-flash-lite | $0.250 | $1.50 | $0.025 | 1M tokens |
| gemini-3-5-flash-lite | $0.300 | $2.50 | $0.030 | 1M tokens |
- Standard paid tier rates shown for text/image/video tokens unless otherwise noted by Google.
- Batch API and Flex pricing are available at lower rates for supported models.
- Context caching storage is billed separately per 1M tokens per hour.
Flash (Paid)
| Model | Input | Output | Cached | Per |
|---|---|---|---|---|
| gemini-2-5-flash | $0.300 | $2.50 | $0.030 | 1M tokens |
| gemini-3-flash-preview | $0.500 | $3.00 | $0.050 | 1M tokens |
| gemini-3-5-flash | $1.50 | $9.00 | $0.150 | 1M tokens |
| gemini-3-6-flash | $1.50 | $7.50 | $0.150 | 1M tokens |
- Standard paid tier rates shown for text/image/video tokens unless otherwise noted by Google.
- Output prices include thinking tokens.
- Batch API and Flex pricing are available at lower rates for supported models.
- Priority inference is priced separately at higher rates.
Pro (Paid)
| Model | Input | Output | Cached | Per |
|---|---|---|---|---|
| gemini-2-5-pro 200K ctx | $2.50 | $15.00 | $0.250 | 1M tokens |
| gemini-3-1-pro-preview 200K ctx | $4.00 | $18.00 | $0.400 | 1M tokens |
- Rates without contextWindow apply to prompts over 200,000 tokens.
- Output prices include thinking tokens.
- Batch API and Flex pricing are available at lower rates for supported models.
- Context caching storage is billed separately per 1M tokens per hour.
Compare Google Gemini API vs Alternatives
Before committing to Google Gemini API, compare pricing with these 3 alternatives in the same category.
What Companies Actually Pay for Google Gemini API
| Model | Input /1M | Output /1M | Blended /1M |
|---|---|---|---|
| google/gemini-3.1-pro-preview | $2.00 | $12.00 | — |
| google/gemini-3-pro-image | $2.00 | $12.00 | — |
| google/gemini-3.5-flash | $1.50 | $9.00 | — |
| google/gemini-2.5-pro | $1.25 | $10.00 | — |
| google/gemini-3.1-flash-image | $0.500 | $3.00 | — |
Google Gemini API Year 1 Total Cost by Company Size
Real deployment costs including licenses, implementation, training, and admin — not just the sticker price.
Solo developer prototyping or building a low-traffic application using the Gemini API Free tier via AI Studio, staying within free-tier rate limits.
A production application processing approximately 10 million input tokens and 5 million output tokens per month using the Flash (Paid) tier via OpenRouter at median provider rates.
A production application processing approximately 10 million input tokens and 5 million output tokens per month using the Pro (Paid) tier at current OpenRouter rates for Gemini 2.5 Pro.
CURRENT TIER DATA
How Google Gemini API Pricing Compares
| Software | Starting Price | Top Price |
|---|---|---|
| Google Gemini API | Free | $18 per million tokens |
| Amazon Bedrock | $0.07 per million tokens | $75 per million tokens |
| Anyscale | Free | $4.9591 per million tokens |
| Baidu ERNIE API | $0.1 per million tokens | $10 per million tokens |
| Cerebras Inference API | $0.1 per million tokens | $6 per million tokens |
| Claude API | $0.1 per million tokens | $75 per million tokens |
Detailed pricing comparisons:
Google Gemini API Contract Terms
Google Gemini API contracts do not auto-renew. Changes require advance notice. These terms are sourced from verified buyer experiences.
How to Negotiate Google Gemini API Pricing
Google Gemini API contracts are negotiable. These 4 tactics are sourced from real buyer experiences and procurement specialists.
The Gemini API Free tier requires no credit card and provides access to all major models via AI Studio. Use this to prototype and measure actual token consumption before committing to any paid volume, giving you precise data to negotiate committed-use discounts.
CURRENT TIER DATAFlash-Lite (via OpenRouter at ~$0.25/1M input) costs approximately 8x less than Pro models (~$2.00/1M input). Benchmark your specific workload across model tiers before defaulting to Pro — many tasks perform acceptably on Flash or Flash-Lite at a fraction of the cost.
OpenRouter pricing dataFor high-volume production workloads, negotiate Committed Use Discounts (CUDs) through Google Cloud's enterprise sales team. CUDs typically require 1-year or 3-year commitments but can yield significant per-token savings over pay-as-you-go rates.
General Google Cloud pricing doctrineOpenRouter routes Gemini models at provider-median blended rates of ~$0.85/1M tokens. Use competing provider pricing (OpenRouter, Vertex AI pricing) as leverage in any enterprise negotiation with Google to justify custom rate discussions.
OpenRouter pricing dataGoogle Gemini API Pricing FAQ
01 How much does the Google Gemini API cost?
Gemini API pricing varies by model. The cheapest option is Gemini 2.5 Flash-Lite at $0.10 per million input tokens and $0.40 per million output tokens. Gemini 2.5 Pro costs $1.25/$10.00 per million tokens (≤200K context). A free tier is available with up to 1,500 requests/day on Flash models via Google AI Studio.
02 Is the Gemini API free?
Yes, Google offers a free tier for the Gemini API through Google AI Studio. The free tier provides access to Flash models with up to 1,500 requests/day and free input/output tokens. Pro models also have a free tier but are rate-limited. For production use, you pay per token on the paid tier with no monthly minimum.
03 Gemini API vs OpenAI API: which is cheaper?
Gemini is generally cheaper than OpenAI for comparable models. Gemini 2.5 Flash at $0.30/$2.50 per million tokens is significantly cheaper than GPT-4o. Gemini 2.5 Pro at $1.25/$10.00 per million tokens undercuts GPT-4o pricing. For budget workloads, Gemini Flash-Lite at $0.10/$0.40 per million tokens has no OpenAI equivalent at that price.
04 What is context caching in the Gemini API?
Context caching lets you cache repeated prompt content (like system instructions or documents) and reuse it across multiple requests. Cached tokens are billed at roughly 90% discount compared to fresh input tokens. This is highly cost-effective for applications that repeatedly process the same large documents or instructions.
05 What is the Batch API discount on Gemini?
The Gemini API Batch API offers a 50% cost reduction on token pricing for asynchronous workloads. Batch requests are processed within 24 hours. This is ideal for offline data processing, bulk classification, or any task that doesn't require real-time responses.
06 Does the Google Gemini API have a free tier?
Yes. The Free tier provides access to Gemini models via AI Studio at no cost, subject to rate limits on requests per minute and per day. It is designed for prototyping and low-volume experimentation, not production-scale workloads.
07 How is the Google Gemini API billed on paid plans?
The Flash-Lite (Paid), Flash (Paid), and Pro (Paid) tiers are all billed on a per-token usage basis with no monthly subscription fee. According to Artificial Analysis data as of April 2026, the provider median across 51 tracked models is $0.56 per 1M input tokens and $2.20 per 1M output tokens, with individual models ranging from near-free (Gemma open models at $0) to premium (Gemini 2.5 Pro at $1.25/$10.00 per 1M input/output tokens).
08 What is the difference between Flash-Lite, Flash, and Pro tiers?
Flash-Lite (Paid) targets the lowest-cost, highest-throughput use cases. Flash (Paid) balances speed and capability for most production workloads. Pro (Paid) is the highest-capability tier suited for complex reasoning tasks. All three are strictly usage-based — there is no monthly minimum or subscription commitment.
09 Can I use Google Gemini API models for free indefinitely?
Yes, through the Free tier. Google provides free access to Gemini models via AI Studio with rate limits, and several Gemma open-weight models are available at $0 per token even on paid infrastructure, according to Artificial Analysis data (April 2026).
10 What models are available through the Gemini API?
The Gemini API offers multiple model tiers: Flash-Lite (the most cost-efficient), Flash (balanced performance and cost), and Pro (highest capability). As of July 2026, specific versions available via OpenRouter include Gemini 2.5 Flash Lite, Gemini 2.5 Flash, Gemini 2.5 Pro, Gemini 3.1 Flash Lite, Gemini 3.1 Flash, Gemini 3.1 Pro Preview, and Gemini 3.5 Flash, among others.
11 When should I upgrade from the free tier to a paid plan?
Upgrade to a paid tier when: (1) you consistently hit free-tier rate limits, (2) you are deploying to production and need higher throughput or SLA guarantees, or (3) you need access to the full Pro model lineup at full context lengths. The free tier is designed for prototyping, not production traffic.
12 Can my free-tier usage data be used to train Google models?
Yes — requests made under the free tier may be used by Google to improve its models. If your use case involves sensitive or proprietary data, you should use the paid API tier, which typically includes data-processing terms that prevent use of your data for model training.
Is this pricing incorrect? — we'll verify and update it.