Budget a Gemini API workload from the model and media it actually processes. Separate text charges, generated media and optional features before comparing free and paid usage.

How Google pricing works

Gemini’s pricing reference separates model entries, input and output types, free and paid tiers, and supported processing modes. Official documentation.

Create separate workload estimates for text answers and generated media. A mixed-media application needs an explicit record of which branch is called and how frequently, rather than a single average text prompt.

Sketch the complete Gemini workflow before entering a token estimate. A document assistant might receive extracted text, a scanned page or a mixture of text and images. An image tool might generate a new asset or edit a supplied one. Keep those paths separate because their input preparation, output settings and charge categories need different evidence.

For each path, record the model identifier, API interface and expected output. Add any optional grounding or other hosted feature explicitly. This prevents a plain-text estimate from quietly becoming the budget for a much broader media application. If a feature’s charge is not represented in the calculator, keep it as an additional unresolved budget item rather than forcing it into an invented token conversion.

PNG diagram separating input modality, output modality, processing mode, cache usage and project billing without literal amounts.

Full price list

Verified model prices
ModelPrice typeUSDUnitTierRegionSource
Gemini 2.5 Computer Use PreviewInput$1.25per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 Computer Use PreviewInput$2.5per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 Computer Use PreviewOutput$10per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 Computer Use PreviewOutput$15per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 FlashBatch Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 FlashBatch Output$1.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 FlashCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 FlashInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 FlashInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 FlashOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 FlashOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 Flash Preview TTSBatch Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash Preview TTSInput$0.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash Preview TTSInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 Flash Preview TTSOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteBatch Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteBatch Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteCached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
Gemini 2.5 Flash-LiteOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 ProBatch Input$0.625per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 ProBatch Input$1.25per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 ProBatch Output$5per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 ProBatch Output$7.5per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 ProCached Input$0.125per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 ProCached Input$0.25per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 ProInput$1.25per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 ProInput$2.5per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 ProInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 2.5 ProOutput$10per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 2.5 ProOutput$15per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 2.5 ProOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3 Flash PreviewBatch Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Flash PreviewBatch Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Flash PreviewCached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Flash PreviewInput$0.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Flash PreviewInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3 Flash PreviewOutput$3per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Flash PreviewOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3 Pro Image (Nano Banana Pro) 🍌Batch Input$1per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Pro Image (Nano Banana Pro) 🍌Batch Output$6per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Pro Image (Nano Banana Pro) 🍌Input$2per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3 Pro Image (Nano Banana Pro) 🍌Output$12per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Image (Nano Banana 2) 🍌Batch Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Image (Nano Banana 2) 🍌Batch Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Image (Nano Banana 2) 🍌Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Image (Nano Banana 2) 🍌Output$3per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) 🍌Batch Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) 🍌Batch Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) 🍌Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) 🍌Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash TTS PreviewBatch Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash TTS PreviewInput$1per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash TTS PreviewInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.1 Flash TTS PreviewOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteBatch Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteBatch Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteOutput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.1 Flash-LiteOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewBatch Input$1per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewBatch Input$2per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewBatch Output$6per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewBatch Output$9per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewCached Input$0.2per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewCached Input$0.4per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewInput$2per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewInput$4per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewOutput$12per 1M tokens<= 200kSee sourceOfficial source ↗
Gemini 3.1 Pro PreviewOutput$18per 1M tokens> 200kSee sourceOfficial source ↗
Gemini 3.5 FlashBatch Input$0.75per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 FlashBatch Output$4.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 FlashCached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 FlashInput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 FlashInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 FlashOutput$9per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 FlashOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteBatch Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteBatch Output$1.25per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 Flash-LiteOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 Live TranslateInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 Live TranslateOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 TranscribeInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 TranscribeOutput$12per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 TranscribeOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 Transcribe LiveInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.5 Transcribe LiveOutput$21per 1M tokensStandardSee sourceOfficial source ↗
Gemini 3.5 Transcribe LiveOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.6 FlashBatch Input$0.375per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.6 FlashBatch Output$1.88per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.6 FlashCached Input$0.075per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.6 FlashInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.6 FlashInput$0.75per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.6 FlashOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.6 FlashOutput$3.75per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.7 FlashBatch Input$0.375per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.7 FlashBatch Output$1.88per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.7 FlashCached Input$0.075per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.7 FlashInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.7 FlashInput$0.75per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.7 FlashOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.7 FlashOutput$3.75per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.8 FlashBatch Input$0.375per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.8 FlashBatch Output$1.88per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.8 FlashCached Input$0.075per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.8 FlashInputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.8 FlashInput$0.75per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini 3.8 FlashOutputFreeper 1M tokensfree tierSee sourceOfficial source ↗
Gemini 3.8 FlashOutput$3.75per 1M tokensthrough December 31, 2026See sourceOfficial source ↗
Gemini Omni FlashInput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini Omni FlashOutput$9per 1M tokensStandardSee sourceOfficial source ↗
Gemini Omni Flash PreviewInput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Gemini Omni Flash PreviewOutput$9per 1M tokensStandardSee sourceOfficial source ↗

Last verified · Source ↗

Read the model entry and the unit together. Compare like with like: a media-output row and a text-output row do not describe equivalent work. Inspect the qualifying mode and the free or paid column before carrying a value into a planning document. Missing data means the relevant fact has not been verified here; it should not be interpreted as a free service.

Follow the source link when a workload depends on a footnote or a particular output setting. Keep the date of the comparison with your saved assumptions. If the application changes its input type or generation mode, revisit the appropriate row instead of continuing to use the original text-only estimate. A useful budget remains connected to the request that actually runs.

Estimate your cost

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

Use the Gemini workload calculator for token-based estimates, and inspect the table’s unit for non-token charges.

Begin with a small representative text request and inspect its returned usage. Keep the selected identifier attached to those observations. Then build separate scenarios for a concise answer, a detailed answer and the relevant conversation context. Label any unmeasured input as an assumption so a colleague can see where the forecast still needs evidence.

For a mixed-media product, estimate each branch independently and combine them only after you understand their units. Keep the branch frequency explicit. A rarely used image feature can still matter to the application budget, while an average across unrelated outputs can hide the behavior you need to control. Preserve the scenarios as a decision record rather than only copying their final totals.

Three worked scenarios

  • Chatbot: compare concise and detailed replies while keeping conversation history explicit.
  • Summarization: identify whether the input is plain text or media that requires separate processing assumptions.
  • Coding assistant: include the code context and generated patch, then measure accepted results.
Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

For a Gemini document assistant, compare a clean text document with a scan that conveys the same information. Evaluate whether the result identifies the requested evidence and whether preprocessing changes the answer. Do not count a fluent description of the document as successful extraction when the application needs specific fields.

For a chatbot, test the same question with a short and a context-rich conversation. Keep relevance and correctness as the acceptance bar while comparing usage. For coding, use a saved repository task and verify the generated change with project checks. If the application needs a repair request, include that work in the evaluation rather than treating the first generated patch as a completed task.

How to reduce cost on Google

The Interactions API supports implicit caching; explicit cache objects belong to the generateContent workflow. Official documentation.

The documented Batch API is an asynchronous generateContent feature. Check interface eligibility before assuming a batch mode can be added to an Interactions request. Official documentation.

Identify which interface the application uses before adding an optimization. The current Gemini guides distinguish Interactions from generateContent for explicit cache objects and batch processing. Make a small proof of the intended request path and result handling before assuming a feature can be enabled with a single option on the existing request.

For repeated material, measure actual cache-related usage and keep the repeated prefix stable during the experiment. Do not infer a cache hit from a faster response alone. For offline batches, assign input identifiers and reconcile individual results before marking the job complete. The lower-cost path is useful only if its lifecycle fits the application and its results still meet the acceptance criteria.

Billing, credits and limits notes

Review billing in the project that owns the application, including the applicable tier and balance behavior. Official documentation.

Read free-tier conditions and project capacity before committing to a deployment budget.

Keep the Cloud project visible in the budget record. Document who owns its billing arrangement, which application uses it and how the workload can be paused. Inspect actual project controls rather than assuming a familiar Google Cloud notification behaves like an enforced stop for every Gemini request.

Separate a prototype’s temporary funding from its ongoing economics. Before enabling a paid continuation, compare current model rates with the measured request mix and decide which optional features are authorized. After representative usage, reconcile the assumptions with the account view. Explain differences in terms of request shape, feature use or account conditions instead of silently replacing the forecast.

Price change history

  1. Google · Gemini 2.5 Computer Use Preview — Amount UsdPrice category: output · USD: 10.000000000 · Unit: per 1M tokens · Tier: > 200k → Price category: output · USD: 15 · Unit: per 1M tokens · Tier: > 200kSource ↗
  2. Google · Gemini 2.5 Computer Use Preview — Amount UsdNot previously recorded → Price category: output · USD: 10 · Unit: per 1M tokens · Tier: <= 200kSource ↗
  3. Google · Gemini 2.5 Computer Use Preview — Amount UsdPrice category: input · USD: 1.250000000 · Unit: per 1M tokens · Tier: > 200k → Price category: input · USD: 2.5 · Unit: per 1M tokens · Tier: > 200kSource ↗

    Parser correction: verified against the official Google pricing table. Restores the documented price for prompts above 200k tokens; not a new provider price change.

  4. Google · Gemini 2.5 Computer Use Preview — Amount UsdNot previously recorded → Price category: input · USD: 1.25 · Unit: per 1M tokens · Tier: <= 200kSource ↗
  5. Google · Gemini Robotics ER 2 Streaming Preview — Amount UsdPrice category: output · USD: 10.000000000 · Unit: per 1M tokens · Tier: > 200k → Price category: output · USD: 15 · Unit: per 1M tokens · Tier: > 200kSource ↗
  6. Google · Gemini Robotics ER 2 Streaming Preview — Amount UsdNot previously recorded → Price category: output · USD: 10 · Unit: per 1M tokens · Tier: <= 200kSource ↗
  7. Google · Gemini Robotics ER 2 Streaming Preview — Amount UsdNot previously recorded → Price category: input · USD: 1.25 · Unit: per 1M tokens · Tier: <= 200kSource ↗
  8. Google · Gemini 2.5 Pro — Amount UsdPrice category: batch_output · USD: 5.000000000 · Unit: per 1M tokens · Tier: > 200k → Price category: batch_output · USD: 7.5 · Unit: per 1M tokens · Tier: > 200kSource ↗
  9. Google · Gemini 2.5 Pro — Amount UsdNot previously recorded → Price category: batch_output · USD: 5 · Unit: per 1M tokens · Tier: <= 200kSource ↗
  10. Google · Gemini 2.5 Pro — Amount UsdPrice category: batch_input · USD: 0.625000000 · Unit: per 1M tokens · Tier: > 200k → Price category: batch_input · USD: 1.25 · Unit: per 1M tokens · Tier: > 200kSource ↗

    Parser correction: verified against the official Google pricing table. Restores the documented price for prompts above 200k tokens; not a new provider price change.

  11. Google · Gemini 2.5 Pro — Amount UsdNot previously recorded → Price category: batch_input · USD: 0.625 · Unit: per 1M tokens · Tier: <= 200kSource ↗
  12. Google · Gemini 2.5 Pro — Amount UsdPrice category: cached_input · USD: 0.125000000 · Unit: per 1M tokens · Tier: > 200k → Price category: cached_input · USD: 0.25 · Unit: per 1M tokens · Tier: > 200kSource ↗

    Parser correction: verified against the official Google pricing table. Restores the documented price for prompts above 200k tokens; not a new provider price change.

Subscribe to the changelog RSS feed

Use a recorded change as a reason to review the affected application paths. Keep the previous model and workload assumptions, then recalculate the same scenarios under the updated source values. A media-model change should not be described as a general improvement in the cost of every Gemini text request.

If a change makes another model attractive, repeat the quality evaluation before switching. Preserve output checks, interface compatibility and data-handling requirements alongside the price comparison. When the official page and a displayed row appear inconsistent, submit the exact model, unit and source link for correction and keep the disputed figure out of an external quote until it is resolved.

Last verified · Source ↗

Frequently asked questions

Is every Gemini input priced like plain text?
No. Read the model’s input and output categories and any media-specific unit. Official documentation.
Does Interactions support explicit cache objects?
The current caching guide directs explicit cache management to generateContent. Official documentation.
Can I use Batch with every interface?
The current Batch guide specifies its supported generateContent workflow. Official documentation.
Will a free-tier row apply to every project?
Confirm both the documented model offer and the billing state of your project.
How should I forecast image generation?
Use the recorded media unit and output settings rather than translating the job into an invented text-token count.

Sources

Last verified · Source ↗