Pricing hub

Gemini API Pricing

Google's Gemini pricing is most useful when you are comparing multimodal and reasoning-heavy workflows against OpenAI or Anthropic at similar token volumes.

Live pricing table

Use this table to compare the published input, cached input, and output rates before you run the calculator.

Input 1,000 / Output 250

ModelInput / 1MCached input / 1MOutput / 1MExample requestSource
Gemini 3 Flash Preview

Google’s preview Flash model for fast multimodal reasoning and agentic workflows.

$0.900No public rate$5.40$0.0023Google Gemini pricing
Gemini 2.5 Pro

Current Gemini Pro model for high-capability multimodal reasoning and coding.

$2.25No public rate$18.0$0.0067Google Gemini pricing
Gemini 2.5 Flash

Fast Google model for balanced cost, latency, and multimodal capability.

$0.540No public rate$4.50$0.0017Google Gemini pricing
Gemini 2.5 Flash Lite

Low-cost Gemini option for high-volume, cost-sensitive workloads.

$0.100No public rate$0.400$0.0002Google Gemini pricing
Gemini 2.0 Flash

Older Gemini Flash tier that still appears in Google’s pricing tables for legacy comparisons.

Google’s current pricing page does not expose a simple public output-token rate for this tier in the same table format.

No public rateNo public rateNo public rateNo public rateMeta pricing docs
Gemini 2.0 Flash Lite

Older ultra-low-cost Gemini tier kept for legacy comparisons and reference.

Google’s current pricing page does not expose a simple public output-token rate for this tier in the same table format.

No public rateNo public rateNo public rateNo public rateMeta pricing docs

How to use this pricing page

  1. 1. Start with real prompts. Estimate the input and output token counts from the workflow you actually run.
  2. 2. Check the example cost. The table shows a request-level estimate so you can compare the published rate card with your usage pattern.
  3. 3. Validate the savings. Open the LLM cost calculator and test a few prompt sizes before you make a budget call.
  4. 4. Compare the model choice. If you need a second opinion, use one of the linked comparison pages below.

Highlights

  • • Good for teams already on Vertex AI.
  • • Helps you compare premium, flash, and lite Gemini tiers.
  • • Use the calculator to estimate the monthly spend of recurring prompt patterns.

Source: Google Vertex AI pricing

FAQ

What is Gemini 3 Flash Preview good for?

It is the current premium Flash-style entry in the shared dataset, so it is the right page for teams comparing quality-first Google workloads.

Should I use this page for budget planning?

Yes. The goal is to give you a stable pricing reference before you turn those rates into per-request and per-month estimates.

Where do I compare it with OpenAI?

Use the comparison pages to see the same input and output profile against GPT-5 or GPT-5 mini.