Pricing hub
Gemini API Pricing
Google's Gemini pricing is most useful when you are comparing multimodal and reasoning-heavy workflows against OpenAI or Anthropic at similar token volumes.
Live pricing table
Use this table to compare the published input, cached input, and output rates before you run the calculator.
Input 1,000 / Output 250
| Model | Input / 1M | Cached input / 1M | Output / 1M | Example request | Source |
|---|---|---|---|---|---|
Gemini 3 Flash Preview Google’s preview Flash model for fast multimodal reasoning and agentic workflows. | $0.900 | No public rate | $5.40 | $0.0023 | Google Gemini pricing |
Gemini 2.5 Pro Current Gemini Pro model for high-capability multimodal reasoning and coding. | $2.25 | No public rate | $18.0 | $0.0067 | Google Gemini pricing |
Gemini 2.5 Flash Fast Google model for balanced cost, latency, and multimodal capability. | $0.540 | No public rate | $4.50 | $0.0017 | Google Gemini pricing |
Gemini 2.5 Flash Lite Low-cost Gemini option for high-volume, cost-sensitive workloads. | $0.100 | No public rate | $0.400 | $0.0002 | Google Gemini pricing |
Gemini 2.0 Flash Older Gemini Flash tier that still appears in Google’s pricing tables for legacy comparisons. Google’s current pricing page does not expose a simple public output-token rate for this tier in the same table format. | No public rate | No public rate | No public rate | No public rate | Meta pricing docs |
Gemini 2.0 Flash Lite Older ultra-low-cost Gemini tier kept for legacy comparisons and reference. Google’s current pricing page does not expose a simple public output-token rate for this tier in the same table format. | No public rate | No public rate | No public rate | No public rate | Meta pricing docs |
How to use this pricing page
- 1. Start with real prompts. Estimate the input and output token counts from the workflow you actually run.
- 2. Check the example cost. The table shows a request-level estimate so you can compare the published rate card with your usage pattern.
- 3. Validate the savings. Open the LLM cost calculator and test a few prompt sizes before you make a budget call.
- 4. Compare the model choice. If you need a second opinion, use one of the linked comparison pages below.
Highlights
- • Good for teams already on Vertex AI.
- • Helps you compare premium, flash, and lite Gemini tiers.
- • Use the calculator to estimate the monthly spend of recurring prompt patterns.
Source: Google Vertex AI pricing
FAQ
What is Gemini 3 Flash Preview good for?
It is the current premium Flash-style entry in the shared dataset, so it is the right page for teams comparing quality-first Google workloads.
Should I use this page for budget planning?
Yes. The goal is to give you a stable pricing reference before you turn those rates into per-request and per-month estimates.
Where do I compare it with OpenAI?
Use the comparison pages to see the same input and output profile against GPT-5 or GPT-5 mini.