Short answer

Google has 8 tracked models in this catalog: 8 with verified first-party prices and 8 with an available OpenRouter route. Use the model pages below for a focused answer and source history.

What are the current Google API prices?

This table sorts tracked Google models by a consistent example workload when a verified first-party price is available. The example uses 1M input tokens and 250K output tokens; it is a comparison aid, not a promise that the models are equivalent.

ModelOfficial in / outOpenRouter in / outContextExample officialExample routedSources
Gemini 2.5 Flash-LiteOfficial API price verified In $0.10Out $0.40 In $0.10Out $0.40 $0.20 $0.20
Gemini 3.1 Flash-LiteOfficial API price verified In $0.25Out $1.50 In $0.25Out $1.50 $0.625 $0.625
Gemini 3.5 Flash-LiteOfficial API price verified In $0.30Out $2.50 In $0.30Out $2.50 $0.925 $0.925
Gemini 2.5 FlashOfficial API price verified In $0.30Out $2.50 In $0.30Out $2.50 1M $0.925 $0.925
Gemini 3.6 FlashOfficial API price verified In $1.50Out $7.50 In $1.50Out $7.50 $3.375 $3.375
Gemini 3.5 FlashOfficial API price verified In $1.50Out $9.00 In $1.50Out $9.00 $3.75 $3.75
Gemini 2.5 ProOfficial API price verified In $1.25Out $10.00 In $1.25Out $10.00 1M $3.75 $3.75
Gemini 3.1 Pro PreviewOfficial API price verified In $2.00Out $12.00 In $2.00Out $12.00 $5.00 $5.00

The first five model pages are also linked here for quick access: Gemini 2.5 Flash-Lite · Gemini 3.1 Flash-Lite · Gemini 3.5 Flash-Lite · Gemini 2.5 Flash · Gemini 3.6 Flash. All prices should be confirmed on the linked source before budgeting.

How should you compare Google models?

Start with the input and output mix of your actual workload, then account for cached input, context-window tiers, tools, retries, and expected monthly runs. A lower output price can still produce a higher monthly bill if the model needs longer responses or more retries.

Use the workload calculator to change token volume, cached-input percentage, and monthly runs. The calculator keeps the official direct route separate from OpenRouter instead of treating a routed price as a first-party quote.

What is the difference between official and OpenRouter Google pricing?

Official pricing refers to a provider-owned API or documentation source. OpenRouter pricing refers to a separate routing and billing channel with its own model ID, availability, verification time, and account terms. The two prices may match, differ, or exist on only one side.

For Google, a missing official hosted price is shown as “No first-party API” rather than being replaced with an OpenRouter number. That distinction is especially important for open-weight or provider-specific routes.

Where can you find each model's source?

Every row links to the provider source used for the official channel and, when available, the exact OpenRouter model record used for the routed channel. The model detail pages include the latest verification date and the number of recorded price-history events.

Prices are checked by the scheduled updater, but a blocked or suspicious source keeps the last verified value instead of silently publishing an unverified change. Read the pricing methodology for the full verification rules.

Google API pricing questions

What is the cheapest Google API model in this catalog?

Gemini 2.5 Flash-Lite has the lowest verified Google example workload cost in this catalog at $0.20 for 1M input tokens and 250K output tokens. A different input-to-output ratio can change the ranking.

Does Google pricing match OpenRouter pricing?

Not necessarily. The official and OpenRouter channels are stored and checked separately for each model. Compare the two columns and follow each source link before making a production decision.

How are Google token prices displayed?

Published standard text-token rates are normalized to USD per one million tokens. Input, cached-input, and output rates remain separate when the source provides them.

Are these Google models interchangeable?

No. Price is only one dimension. Context limits, quality, latency, tools, availability, safety behavior, and provider terms still need to be tested for the intended workload.