Exact model · kimi-k3

Kimi K3 API pricing and monthly cost by provider

Compare the exact kimi-k3 route across owner-direct, router, and serverless channels while keeping provider configuration visible.

Exact model · kimi-k3

Kimi K3 monthly cost comparison

Compare the exact kimi-k3 route across owner-direct, router, and serverless channels while keeping provider configuration visible.

No complete comparison
Versioned service or route ID; family-only names are excluded.
A route without a reviewed cache meter is excluded only when this value is above zero.
No complete comparison. No provider has complete, fresh evidence for this exact workload. APIDir does not estimate missing meters.

USD-normalized monthly cost for the entered workload; every meter is per 1 million tokens.
Rank statusProviderInput priceCache-read priceOutput priceMonthly costExact configurationSources and freshness
—Evidence is aging or staleFireworks AIaggregator$3 / 1M input tokens$0.3 / 1M cached input tokens$15 / 1M output tokensUnavailableEvidence is aging or stalefireworks/kimi-k3 · serverless standard
  • Input · checked · fresh through staleFirst-party source
  • Cached input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
—Evidence is aging or staleKimi API Platformofficial$3 / 1M input tokens$0.3 / 1M cached input tokens$15 / 1M output tokensUnavailableEvidence is aging or stalekimi-k3 · owner-direct pay-as-you-go
  • Input · checked · fresh through staleFirst-party source
  • Cached input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
—Evidence is aging or staleNovita AIaggregator$3 / 1M input tokens$0.3 / 1M cached input tokens$15 / 1M output tokensUnavailableEvidence is aging or stalemoonshotai/kimi-k3 · serverless · chat completions
  • Input · checked · fresh through staleFirst-party source
  • Cached input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
—Evidence is aging or staleOpenRouterrouter$3 / 1M input tokens$0.3 / 1M cached input tokens$15 / 1M output tokensUnavailableEvidence is aging or stalemoonshotai/kimi-k3 · standard · chat completions
  • Input · checked · fresh through staleFirst-party source
  • Cached input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
—Evidence is aging or staleTogether AIaggregator$3 / 1M input tokensNot published$15 / 1M output tokensUnavailableEvidence is aging or stalemoonshotai/Kimi-K3 · serverless · chat completions
  • Input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
Missing and excluded providers (6)
  • Fireworks AI: Evidence is aging or stale.
  • Kimi API Platform: Evidence is aging or stale.
  • Novita AI: Evidence is aging or stale.
  • OpenRouter: Evidence is aging or stale.
  • Together AI: Evidence is aging or stale.
  • DeepInfra: No accepted exact Kimi K3 route and price observation was found in this review.

Strict boundary: Fireworks standard serverless is used; priority, fast, and US-region premiums are excluded. Together has no reviewed cached-input meter and is excluded when cached usage is above zero. The comparison does not score throughput, latency, limits, reliability, support, or data terms.

Embeddable result

Put this exact-model comparison on your site

The SVG card is cookie-free and links back to the same source-auditable workload. It shows a rank only while at least three routes qualify.

<script async referrerpolicy="no-referrer" src="https://apidir.dev/embed/model-cost.js?model=kimi-api&input=100000000&cached=0&output=20000000"></script>

One exact AI model compared across three provider routes with separate token meters and an incomplete evidence route
AI concept diagramConcept diagram: exact identity and complete meters decide eligibility. An incomplete route stays visible as a gap and cannot become a winner.

Exact-model alert · controlled rollout

Join the Kimi K3 price alert waitlist.

Confirm your address to join the waitlist. Delivery remains paused until sender, suppression, and evidence checks are ready; ambiguous or stale changes stay quarantined.

Evidence boundary

What the monthly comparison proves

Fireworks standard serverless is used; priority, fast, and US-region premiums are excluded. Together has no reviewed cached-input meter and is excluded when cached usage is above zero. The comparison does not score throughput, latency, limits, reliability, support, or data terms.

A row qualifies only when the exact model route, channel, region, mode, tier, and input/output meters match the reviewed definition. Cached input becomes mandatory when your cached-input volume is above zero.

Questions developers ask

Which exact Kimi K3 route is compared?

This page is locked to kimi-k3. Family-only names, near matches, different model versions, and unreviewed aliases are excluded.

When is a lowest-cost label shown?

Only when at least three unique providers have complete fresh evidence for every token meter used by the submitted workload. Equal monthly costs share a rank.

What is excluded from monthly cost?

Retries, failed requests, networking, storage, taxes, support, dedicated capacity, latency, rate limits, and provider-specific service quality are outside this token-meter estimate.