Moonshot AI · text · family comparison

Kimi API pricing by model and provider

Compare Kimi API pricing across provider routes without mixing Kimi versions, token meters, context tiers, or service channels.

Current answer: 16 accepted Kimi API observations; one current example is $0.3 / 1M cached input tokens, checked Aug 22, 2026. Open the evidence source, then compare an exact matching configuration below.

Exact-model monthly workload

Kimi K3 price ranking status

No complete current comparison

Default workload: 100M uncached input + 20M output tokens. No provider has complete, fresh evidence for this exact workload. APIDir does not estimate missing meters.

Why there is no winner

No provider has complete, fresh evidence for this exact workload. APIDir does not estimate missing meters.

Enter your own token volume

Model-level provider comparison

Kimi API price ranking status

Reviewed scopes, no current winner

Current answer: No current cheapest provider is published. The reviewed scopes do not have enough fresh, like-for-like offers to name a winner.

Provider routes with evidence
5
Accepted observations
16
Billing meters
1M cached input tokens, 1M input tokens, 1M output tokens
Latest accepted check
Aug 22, 2026

Kimi K3 API input token pricing

Kimi K3 input token price ranking

Not currently rankable
Reviewed comparison scope
Kimi K3 · shared API route · global · USD per 1M input tokens
Publication gate
At least 3 unique providers with accepted, fresh, reviewed evidence

Why there is no winner

Ranking withheld: 0 of 3 required providers have fresh, matching evidence.

The three providers identify Kimi K3 and the input token meter. Tier names and service features remain visible and are not treated as equivalent quality, speed, or reliability.

Kimi K3 API output token pricing

Kimi K3 output token price ranking

Not currently rankable
Reviewed comparison scope
Kimi K3 · shared API route · global · USD per 1M output tokens
Publication gate
At least 3 unique providers with accepted, fresh, reviewed evidence

Why there is no winner

Ranking withheld: 0 of 3 required providers have fresh, matching evidence.

The order covers the listed output token meter only. Context, cache, throughput, quotas, endpoint behavior, and support remain separate buying criteria.

Uncollapsed source records

All accepted offer evidence

These rows remain separate until their offered model ID, route, meter, tier, resolution, duration, audio, and region are proven comparable.

Kimi API provider offers · exact configuration scope
ProviderModelConfigurationRecorded priceEvidence
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless standard · not-applicable · no audio$0.3 / 1M cached input tokensstaleProvider sourceChecked Aug 22, 2026
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless standard · not-applicable · no audio$3 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio$3 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio$15 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026
Fireworks AIaggregatorKimi APImodel: fireworks/kimi-k3 · chat-completions · serverless standard · not-applicable · no audio$15 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026
Kimi API PlatformofficialKimi APImodel: kimi-k3 · chat-completions · Kimi K3 pay-as-you-go · not-applicable · no audio$0.3 / 1M cached input tokensstaleProvider sourceChecked Aug 22, 2026
Kimi API PlatformofficialKimi APImodel: kimi-k3 · chat-completions · Kimi K3 pay-as-you-go · not-applicable · no audio$3 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
Kimi API PlatformofficialKimi APImodel: kimi-k3 · chat-completions · Kimi K3 pay-as-you-go · not-applicable · no audio$15 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026
Novita AIaggregatorKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 serverless · not-applicable · no audio$0.3 / 1M cached input tokensstaleProvider sourceChecked Aug 22, 2026
Novita AIaggregatorKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 serverless · not-applicable · no audio$3 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
Novita AIaggregatorKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 serverless · not-applicable · no audio$15 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026
OpenRouterrouterKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio$0.3 / 1M cached input tokensstaleProvider sourceChecked Aug 22, 2026
OpenRouterrouterKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio$3 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
OpenRouterrouterKimi APImodel: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio$15 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026
Together AIaggregatorKimi APImodel: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio$3 / 1M input tokensstaleProvider sourceChecked Aug 22, 2026
Together AIaggregatorKimi APImodel: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio$15 / 1M output tokensstaleProvider sourceChecked Aug 22, 2026

Checked evidence, not a ranking: fresh rows were checked against the linked first-party source on the shown date. APIDir does not name a cheapest provider unless exact configurations and units match. Recheck before purchasing.

Developer resources

Provider documentation and pricing links

Compare integration documentation beside the exact first-party source used for each accepted observation.

Provider documentation, pricing, and accepted evidence for this model
ProviderChannelDeveloper resourcesAccepted evidence
Fireworks AIaggregatorDocumentationPricing page5 observationsEvidence source 1
Kimi API PlatformofficialDocumentationPricing page3 observationsEvidence source 1
Novita AIaggregatorDocumentationPricing page3 observationsEvidence source 1
OpenRouterrouterDocumentationPricing page3 observationsEvidence source 1
Together AIaggregatorDocumentationPricing page2 observationsEvidence source 1

Route-specific buying guide

How to budget Kimi K3 on OpenRouter

Route ID: moonshotai/kimi-k3

The accepted rows describe OpenRouter's Kimi K3 route. They are not Moonshot's direct developer prices and do not establish a universal cheapest Kimi provider.

OpenRouter reported a 1,048,576-token context window when checked. Context capacity is a ceiling, not a budget target; retrieval and conversation growth can turn input tokens into the dominant cost.

Cost formula

Estimate one request as uncached input tokens × the input rate, cached input tokens × the cached-input rate, plus output tokens × the output rate. Divide every token count by one million before applying APIDir's displayed rates.

Costs that move the bill

  • Long documents, retrieval payloads, conversation history, and agent memory
  • Completion length and repeated reasoning, coding, browsing, or tool-use loops
  • Cache coverage, cache invalidation frequency, and any route-specific write charge
  • Retries, timeouts, fallback models, rate limits, and production support needs

Checks before purchase

  • Keep the moonshotai/kimi-k3 route ID and the checked date beside every cost result.
  • Benchmark short, median, and long prompts because a single average hides tail cost.
  • Cap output and agent steps deliberately, then measure successful-task cost rather than request cost.
  • Treat OpenRouter and Moonshot direct access as different offers until their commercial boundaries match.

Source checked Aug 22, 2026: OpenRouter public model API. Recheck the live route before committing spend.

Estimate this route with the LLM workload calculator →

Compare another model family

Decision guide

How to compare Kimi API providers

Kimi is a model family, so a useful provider comparison must name the exact release. The reviewed ranking uses Kimi K3 across OpenRouter, Together AI, and Fireworks AI. Each provider's route spelling remains visible even when the underlying release and token meter match.

Equal token rates do not prove equal service. Context support, cache behavior, throughput, quotas, fallback, endpoint compatibility, regional access, data terms, and support remain separate buying criteria. Use the ranking as a price-meter result and benchmark the service characteristics yourself.

What changes Kimi API cost

  • Exact Kimi release, input-output ratio, context, and cache use
  • Router or aggregator tier, throughput, rate limits, and any dedicated option
  • Tool calls, retries, data retention, regional access, and support

Integration checks before production

  • Pin Kimi K3's full provider route ID and reject silent fallback
  • Set context, output, timeout, and retry budgets at the caller boundary
  • Log token counts, route, terminal status, retry count, and request ID

Workloads this comparison can inform

Long-context assistants

Verify context support and cache meters before extending a simple token estimate.

Batch language processing

Measure real output ratios and concurrency instead of ranking input price alone.

Multi-provider routing

Test fallback identity and error semantics even when the headline rates tie.

How to read the price result

Three reviewed Kimi K3 routes currently publish the same standard input and output rates. APIDir reports the tie for those meters only; it is not a quality, speed, uptime, or support ranking.

Related model price comparisons

Use this evidence safely

A recorded price is not a complete production quote. Open the source, reproduce the exact configuration, and include retries, failed work, storage, egress, support, rate limits, and downstream processing in total cost.

Questions developers ask

What makes two Kimi offers comparable?

The exact model and endpoint, provider route, input and output meters, cache policy, context-length price band, tools, region, and service tier must match.

Does APIDir name a cheapest provider on this page?

Only when at least three explicitly reviewed provider offers are fresh, publication-permitted, and within the same published comparison scope. Stale or mismatched evidence cannot support that claim.

How should I confirm a budget?

Open the linked provider source, verify the configuration, then calculate your own successful and failed workload volume.

APIDir Delta · controlled rollout

Join the source-linked pricing update waitlist.

Confirm your address to join the waitlist. Editorial delivery starts only after sender, suppression, and evidence checks are active.