Moonshot AI · text · family comparison

Kimi API pricing by model and provider

Compare Kimi API pricing across provider routes without mixing Kimi versions, token meters, context tiers, or service channels.

Model-level provider comparison

Kimi API price ranking status

2 published price rankings

Current answer: APIDir currently publishes 2 comparable price rankings. Keep each meter and configuration separate before choosing a route.

Provider routes with evidence
3
Accepted observations
7
Billing meters
1M cached input tokens, 1M input tokens, 1M output tokens
Latest accepted check
Aug 22, 2026

Kimi K3 API input token pricing

Kimi K3 input token price ranking

Published ranking
Reviewed comparison scope
Kimi K3 · shared API route · global · USD per 1M input tokens
Publication gate
At least 3 unique providers with accepted, fresh, reviewed evidence
Kimi K3 · shared API route · global · USD per 1M input tokens
Price rankProviderRoute and tierPublished meterEvidence
1TieFireworks AIaggregator · serverless-standard$3 / 1M input tokensfreshFirst-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026
1TieOpenRouterrouter · Kimi K3 standard$3 / 1M input tokensfreshFirst-party sourceChecked Aug 22, 2026Fresh until Aug 25, 2026
1TieTogether AIaggregator · serverless$3 / 1M input tokensfreshFirst-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026

Price meter only: The three providers identify Kimi K3 and the input token meter. Tier names and service features remain visible and are not treated as equivalent quality, speed, or reliability. Equal prices share a price rank; provider name and offer ID determine only their stable display order.

Kimi K3 API output token pricing

Kimi K3 output token price ranking

Published ranking
Reviewed comparison scope
Kimi K3 · shared API route · global · USD per 1M output tokens
Publication gate
At least 3 unique providers with accepted, fresh, reviewed evidence
Kimi K3 · shared API route · global · USD per 1M output tokens
Price rankProviderRoute and tierPublished meterEvidence
1TieFireworks AIaggregator · serverless-standard$15 / 1M output tokensfreshFirst-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026
1TieOpenRouterrouter · Kimi K3 standard$15 / 1M output tokensfreshFirst-party sourceChecked Aug 22, 2026Fresh until Aug 25, 2026
1TieTogether AIaggregator · serverless$15 / 1M output tokensfreshFirst-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026

Price meter only: The order covers the listed output token meter only. Context, cache, throughput, quotas, endpoint behavior, and support remain separate buying criteria. Equal prices share a price rank; provider name and offer ID determine only their stable display order.

Uncollapsed source records

All accepted offer evidence

These rows remain separate until their offered model ID, route, meter, tier, resolution, duration, audio, and region are proven comparable.

Kimi API provider offers · exact configuration scope
ProviderModelConfigurationRecorded priceEvidence
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio$3 / 1M input tokensfreshProvider sourceChecked Aug 22, 2026
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio$15 / 1M output tokensfreshProvider sourceChecked Aug 22, 2026
OpenRouterrouterKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio$0.3 / 1M cached input tokensfreshProvider sourceChecked Aug 22, 2026
OpenRouterrouterKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio$3 / 1M input tokensfreshProvider sourceChecked Aug 22, 2026
OpenRouterrouterKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio$15 / 1M output tokensfreshProvider sourceChecked Aug 22, 2026
Together AIaggregatorKimi APImodel: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio$3 / 1M input tokensfreshProvider sourceChecked Aug 22, 2026
Together AIaggregatorKimi APImodel: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio$15 / 1M output tokensfreshProvider sourceChecked Aug 22, 2026

Checked evidence, not a ranking: fresh rows were checked against the linked first-party source on the shown date. APIDir does not name a cheapest provider unless exact configurations and units match. Recheck before purchasing.

Developer resources

Provider documentation and pricing links

Compare integration documentation beside the exact first-party source used for each accepted observation.

Provider documentation, pricing, and accepted evidence for this model
ProviderChannelDeveloper resourcesAccepted evidence
Fireworks AIaggregatorDocumentationPricing page2 observationsEvidence source 1
OpenRouterrouterDocumentationPricing page3 observationsEvidence source 1
Together AIaggregatorDocumentationPricing page2 observationsEvidence source 1

Route-specific buying guide

How to budget Kimi K3 on OpenRouter

Route ID: moonshotai/kimi-k3

The accepted rows describe OpenRouter's Kimi K3 route. They are not Moonshot's direct developer prices and do not establish a universal cheapest Kimi provider.

OpenRouter reported a 1,048,576-token context window when checked. Context capacity is a ceiling, not a budget target; retrieval and conversation growth can turn input tokens into the dominant cost.

Cost formula

Estimate one request as uncached input tokens × the input rate, cached input tokens × the cached-input rate, plus output tokens × the output rate. Divide every token count by one million before applying APIDir's displayed rates.

Costs that move the bill

  • Long documents, retrieval payloads, conversation history, and agent memory
  • Completion length and repeated reasoning, coding, browsing, or tool-use loops
  • Cache coverage, cache invalidation frequency, and any route-specific write charge
  • Retries, timeouts, fallback models, rate limits, and production support needs

Checks before purchase

  • Keep the moonshotai/kimi-k3 route ID and the checked date beside every cost result.
  • Benchmark short, median, and long prompts because a single average hides tail cost.
  • Cap output and agent steps deliberately, then measure successful-task cost rather than request cost.
  • Treat OpenRouter and Moonshot direct access as different offers until their commercial boundaries match.

Source checked Aug 22, 2026: OpenRouter public model API. Recheck the live route before committing spend.

Estimate this route with the LLM workload calculator →

Compare another model family

Decision guide

How to compare Kimi API API providers

Kimi is a model family, so a useful provider comparison must name the exact release. The reviewed ranking uses Kimi K3 across OpenRouter, Together AI, and Fireworks AI. Each provider's route spelling remains visible even when the underlying release and token meter match.

Equal token rates do not prove equal service. Context support, cache behavior, throughput, quotas, fallback, endpoint compatibility, regional access, data terms, and support remain separate buying criteria. Use the ranking as a price-meter result and benchmark the service characteristics yourself.

What changes the Kimi API API cost

  • Exact Kimi release, input-output ratio, context, and cache use
  • Router or aggregator tier, throughput, rate limits, and any dedicated option
  • Tool calls, retries, data retention, regional access, and support

Integration checks before production

  • Pin Kimi K3's full provider route ID and reject silent fallback
  • Set context, output, timeout, and retry budgets at the caller boundary
  • Log token counts, route, terminal status, retry count, and request ID

Workloads this comparison can inform

Long-context assistants

Verify context support and cache meters before extending a simple token estimate.

Batch language processing

Measure real output ratios and concurrency instead of ranking input price alone.

Multi-provider routing

Test fallback identity and error semantics even when the headline rates tie.

How to read the price result

Three reviewed Kimi K3 routes currently publish the same standard input and output rates. APIDir reports the tie for those meters only; it is not a quality, speed, uptime, or support ranking.

Related model price comparisons

Use this evidence safely

A recorded price is not a complete production quote. Open the source, reproduce the exact configuration, and include retries, failed work, storage, egress, support, rate limits, and downstream processing in total cost.

Questions developers ask

What makes two Kimi offers comparable?

The exact model and endpoint, provider route, input and output meters, cache policy, context-length price band, tools, region, and service tier must match.

Does APIDir name a cheapest provider on this page?

Only when at least three explicitly reviewed provider offers are fresh, publication-permitted, and within the same published comparison scope. Stale or mismatched evidence cannot support that claim.

How should I confirm a budget?

Open the linked provider source, verify the configuration, then calculate your own successful and failed workload volume.

APIDir Delta · controlled rollout

Reserve source-linked price research.

Confirm your address to join the rollout. Editorial delivery starts only after sender, suppression, and evidence checks are active.