Google · multimodal · family comparison

Gemini API pricing by model and provider

Compare Gemini API routes by exact model, service channel, standard or batch tier, token meter, and first-party source.

Current answer: 4 accepted Gemini API observations; one current example is $0.75 / 1M input tokens, checked Aug 22, 2026. Open the evidence source, then compare an exact matching configuration below.

Model-level provider comparison

Gemini API price ranking status

Evidence tracked, not yet rankable

Current answer: No current cheapest provider is published. Accepted evidence exists, but it does not yet form a named, like-for-like comparison group.

Provider routes with evidence
2
Accepted observations
4
Billing meters
1M input tokens, 1M output tokens
Latest accepted check
Aug 22, 2026

Why there is no winner

No reviewed comparison scope exists for this model yet. Accepted evidence can still be inspected below, but APIDir will not sort unlike offers or infer a cheapest provider.

Read the publication gate

Uncollapsed source records

All accepted offer evidence

These rows remain separate until their offered model ID, route, meter, tier, resolution, duration, audio, and region are proven comparable.

Gemini API provider offers · exact configuration scope
ProviderModelConfigurationRecorded priceEvidence
Google Ai Studio APIofficialGemini APImodel: gemini-3.7-flash · generate-content · standard-introductory · not-applicable · no audio$0.75 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
Google Ai Studio APIofficialGemini APImodel: gemini-3.7-flash · generate-content · standard-introductory · not-applicable · no audio$3.75 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026
OpenRouterrouterGemini APImodel: google/gemini-3.7-flash · chat-completions · batch-route · not-applicable · no audio$0.375 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
OpenRouterrouterGemini APImodel: google/gemini-3.7-flash · chat-completions · batch-route · not-applicable · no audio$1.875 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026

Checked evidence, not a ranking: fresh rows were checked against the linked first-party source on the shown date. APIDir does not name a cheapest provider unless exact configurations and units match. Recheck before purchasing.

Developer resources

Provider documentation and pricing links

Compare integration documentation beside the exact first-party source used for each accepted observation.

Provider documentation, pricing, and accepted evidence for this model
ProviderChannelDeveloper resourcesAccepted evidence
Google Ai Studio APIofficialDocumentationPricing page2 observationsEvidence source 1
OpenRouterrouterDocumentationPricing page2 observationsEvidence source 1

Decision guide

How to compare Gemini API providers

Gemini API pricing varies by exact model, input and output modality, context, caching, grounding, and service tier. APIDir currently records Gemini 3.7 Flash through Google AI Studio and an OpenRouter batch route. The lower batch number is not labelled cheaper than standard service because delivery semantics differ.

Multimodal requests require more than a text-token estimate. Images, audio, video, cached context, search grounding, and generated media can have separate meters. Confirm which service accepts the required modality and preserve the route and tier in your cost logs.

What changes Gemini API cost

  • Exact Gemini model, standard or batch tier, context, and input-output ratio
  • Image, audio, video, cache, grounding, and tool-use meters
  • Regional service, quotas, retry behavior, storage, and data-governance requirements

Integration checks before production

  • Pin the service channel and model ID together
  • Validate modality limits and payload size before sending a billable request
  • Track tier, token usage, media units, retries, and provider request ID

Workloads this comparison can inform

Multimodal extraction

Model text and media units separately; a text-only comparison understates the workload.

Asynchronous batch processing

A batch route can reduce the meter when latency and completion guarantees fit the job.

Interactive applications

Use standard-service evidence and measure latency, quotas, and fallback behavior.

How to read the price result

The current direct row is standard introductory service while the router row is explicitly batch. Their prices are displayed side by side as different routes, not sorted into one winner.

Related model price comparisons

Use this evidence safely

A recorded price is not a complete production quote. Open the source, reproduce the exact configuration, and include retries, failed work, storage, egress, support, rate limits, and downstream processing in total cost.

Questions developers ask

What makes two Gemini offers comparable?

The model version, generation mode, tier, resolution, duration, audio setting, region, and billing unit must match.

Does APIDir name a cheapest provider on this page?

Only when at least three explicitly reviewed provider offers are fresh, publication-permitted, and within the same published comparison scope. Stale or mismatched evidence cannot support that claim.

How should I confirm a budget?

Open the linked provider source, verify the configuration, then calculate your own successful and failed workload volume.

APIDir Delta · controlled rollout

Join the source-linked pricing update waitlist.

Confirm your address to join the waitlist. Editorial delivery starts only after sender, suppression, and evidence checks are active.