Google Gemini Developer API

Google Gemini Developer API supplies hosted language models with token-based billing. The calculator keeps model rates and limits in a versioned pricing snapshot.

Categories
LLM
Models and endpoints
3
Guide reviewed
Sep 24, 2026

Models, translated into money

Per million tokens. See what your budget actually buys.

TaskBudget · USD
Compare3 models
Task
Budget USD

Gemini 3.5 Flash-Lite

Google

What $10 gets you

About 5,405chats

Where it fits

Good fit
ChatSupportDocuments
Limited fit
Coding
Calculate with Gemini 3.5 Flash-LiteUse in calculator

ByeTokens

Gemini 3.5 Flash-Lite


  • Your budget$10
  • Buys about5,405 chats
  • WorkloadTypical chat2,000 input + 500 output tokens per request

  • Price examples
    100 chats$0.1850
  • 1,000 chats$1.85
  • 10,000 chats$18.50

Catalog 2026-09-26-7e5c5fac
Each line is a separate example.

How we calculated thisVerified Sep 26
Typical chat assumption2,000 input · 500 output tokens per request

Official pricing

google-gemini-3-5-flash-lite-standard · Current · Standard billing

Standard text
$0.30 input · $2.50 output / 1M tokens

Verified Sep 26, 2026

Model details

Model ID
gemini-3.5-flash-lite
Tool calling
true
Structured output
true
context Window
1,048,576
max Output Tokens
65,536
modality
text
processing Tier
standard
thinking Tokens Included
Yes

Task evidence

General chat: 84 / 100 · good fit

Support chat: 84 / 100 · good fit

Coding help: 67 / 100 · limited fit

Document work: 82 / 100 · good fit

Research agent: No current score

Catalog 2026-09-26-7e5c5fac

Gemini 3.8 Flash

Google

What $10 gets you

About 2,962chats

Where it fits

Strong fit
ChatSupportDocumentsResearch
Good fit
Coding
Calculate with Gemini 3.8 FlashUse in calculator

ByeTokens

Gemini 3.8 Flash


  • Your budget$10
  • Buys about2,962 chats
  • WorkloadTypical chat2,000 input + 500 output tokens per request

  • Price examples
    100 chats$0.3375
  • 1,000 chats$3.38
  • 10,000 chats$33.75

Catalog 2026-09-26-7e5c5fac
Each line is a separate example.

How we calculated thisVerified Sep 26
Typical chat assumption2,000 input · 500 output tokens per request

Official pricing

google-gemini-3-8-flash-standard-2026 · Until Dec 31, 2026 · Standard billing

Standard text
$0.75 input · $3.75 output / 1M tokens

Verified Sep 26, 2026

google-gemini-3-8-flash-standard-2027 · Starts Jan 1, 2027 · Standard billing

Standard text
$1.50 input · $7.50 output / 1M tokens

Verified Sep 26, 2026

Model details

Model ID
gemini-3.8-flash
Tool calling
true
Structured output
true
context Window
1,048,576
max Output Tokens
65,536
modality
text
processing Tier
standard
thinking Tokens Included
Yes

Task evidence

General chat: 98 / 100 · strong fit

Support chat: 98 / 100 · strong fit

Coding help: 88 / 100 · good fit

Document work: 98 / 100 · strong fit

Research agent: 91 / 100 · strong fit

Catalog 2026-09-26-7e5c5fac

Gemini 3.1 Pro Preview

Google

What $10 gets you

About 1,000chats

Where it fits

Strong fit
ChatSupport
Good fit
CodingDocumentsResearch
Calculate with Gemini 3.1 Pro PreviewUse in calculator

ByeTokens

Gemini 3.1 Pro Preview


  • Your budget$10
  • Buys about1,000 chats
  • WorkloadTypical chat2,000 input + 500 output tokens per request

  • Price examples
    100 chats$1.00
  • 1,000 chats$10.00
  • 10,000 chats$100.00

Catalog 2026-09-26-7e5c5fac
Each line is a separate example.

How we calculated thisVerified Sep 26
Typical chat assumption2,000 input · 500 output tokens per request

Official pricing

google-gemini-3-1-pro-preview-standard · Current · Standard billing

Standard text
$2.00 input · $12.00 output / 1M tokens
Above 200,000 input tokens per request
$4.00 input · $18.00 output / 1M tokens

Verified Sep 26, 2026

Model details

Model ID
gemini-3.1-pro-preview
Tool calling
true
Structured output
true
context Window
1,048,576
max Output Tokens
65,536
modality
text
processing Tier
standard
thinking Tokens Included
Yes

Task evidence

General chat: 96 / 100 · strong fit

Support chat: 96 / 100 · strong fit

Coding help: 73 / 100 · good fit

Document work: 79 / 100 · good fit

Research agent: 76 / 100 · good fit

Catalog 2026-09-26-7e5c5fac

Google Gemini Developer API is included as a hosted language-model provider. The catalog focuses on paid API usage for an application, not consumer Gemini plans and not bundled cloud commitments. This keeps the estimate tied to an exact endpoint and a reproducible input/output workload.

Where Gemini fits

The catalogued Gemini model can cover chat, generation, extraction, and research-oriented language calls. A complete AI product may also require web search, page retrieval, image generation, or video generation. Those services are modeled separately because their billing units and compatibility rules differ from language tokens.

The scenario builder translates product activity into billable work. For example, monthly users and messages per user produce chat calls, while research tasks can produce multiple language calls plus search and scraping demand. Exact mode is available when the team already knows monthly request and token totals.

Token pricing without hidden averaging

Input and output tokens are priced independently. Chat and research remain separate rows through context-band selection, even when both use the same model. This is important because a large prompt can be billed or constrained differently from a small one. The engine calculates each row with decimal arithmetic and rounds only for display.

The live offer section above supplies the exact model ID, current public plan, limits, capabilities, verification date, and official source links. These details are rendered from the current catalog snapshot. The article explains the decision model but does not hard-code prices that could become stale.

Only the billing path described by the active plan is included. Free tiers, batch processing, caching, priority modes, experimental endpoints, and negotiated agreements are excluded unless represented by a separate verified offer. An estimate should not assume that a free allowance repeats every month unless the user explicitly enables a catalogued recurring allowance.

Compatibility also matters. The calculator checks minimum context, tool-calling support, structured output, and model limits. Unknown support fails closed for required capabilities. This can remove an offer from a recommendation even when a price is known.

Use the verified offer

Select the offer above to open the calculator with Chat enabled and the model locked. The rest of the stack remains comparable, so search, scraping, image, and video providers can still change independently. This is useful for answering a narrow question such as “What would this application cost if Gemini handled every language call?”

The estimate excludes taxes, infrastructure, storage, retries, and development costs. Check the displayed source and verification date before using the result for a production decision.