Moonshot AI · text · family comparison
Kimi API pricing by model and provider
Compare Kimi API pricing across provider routes without mixing Kimi versions, token meters, context tiers, or service channels.
Exact-model monthly workload
Kimi K3 price ranking status
Default workload: 100M uncached input + 20M output tokens. No provider has complete, fresh evidence for this exact workload. APIDir does not estimate missing meters.
Why there is no winner
No provider has complete, fresh evidence for this exact workload. APIDir does not estimate missing meters.
Model-level provider comparison
Kimi API price ranking status
Current answer: No current cheapest provider is published. The reviewed scopes do not have enough fresh, like-for-like offers to name a winner.
- Provider routes with evidence
- 5
- Accepted observations
- 16
- Billing meters
- 1M cached input tokens, 1M input tokens, 1M output tokens
- Latest accepted check
- Aug 22, 2026
Kimi K3 API input token pricing
Kimi K3 input token price ranking
- Reviewed comparison scope
- Kimi K3 · shared API route · global · USD per 1M input tokens
- Publication gate
- At least 3 unique providers with accepted, fresh, reviewed evidence
Why there is no winner
Ranking withheld: 0 of 3 required providers have fresh, matching evidence.
The three providers identify Kimi K3 and the input token meter. Tier names and service features remain visible and are not treated as equivalent quality, speed, or reliability.
Kimi K3 API output token pricing
Kimi K3 output token price ranking
- Reviewed comparison scope
- Kimi K3 · shared API route · global · USD per 1M output tokens
- Publication gate
- At least 3 unique providers with accepted, fresh, reviewed evidence
Why there is no winner
Ranking withheld: 0 of 3 required providers have fresh, matching evidence.
The order covers the listed output token meter only. Context, cache, throughput, quotas, endpoint behavior, and support remain separate buying criteria.
Uncollapsed source records
All accepted offer evidence
These rows remain separate until their offered model ID, route, meter, tier, resolution, duration, audio, and region are proven comparable.
| Provider | Model | Configuration | Recorded price | Evidence |
|---|---|---|---|---|
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless standard · not-applicable · no audio | $0.3 / 1M cached input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless standard · not-applicable · no audio | $3 / 1M input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio | $3 / 1M input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio | $15 / 1M output tokensstale | Provider sourceChecked Aug 22, 2026 |
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless standard · not-applicable · no audio | $15 / 1M output tokensstale | Provider sourceChecked Aug 22, 2026 |
| Kimi API Platformofficial | Kimi API | model: kimi-k3 · chat-completions · Kimi K3 pay-as-you-go · not-applicable · no audio | $0.3 / 1M cached input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Kimi API Platformofficial | Kimi API | model: kimi-k3 · chat-completions · Kimi K3 pay-as-you-go · not-applicable · no audio | $3 / 1M input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Kimi API Platformofficial | Kimi API | model: kimi-k3 · chat-completions · Kimi K3 pay-as-you-go · not-applicable · no audio | $15 / 1M output tokensstale | Provider sourceChecked Aug 22, 2026 |
| Novita AIaggregator | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 serverless · not-applicable · no audio | $0.3 / 1M cached input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Novita AIaggregator | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 serverless · not-applicable · no audio | $3 / 1M input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Novita AIaggregator | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 serverless · not-applicable · no audio | $15 / 1M output tokensstale | Provider sourceChecked Aug 22, 2026 |
| OpenRouterrouter | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio | $0.3 / 1M cached input tokensstale | Provider sourceChecked Aug 22, 2026 |
| OpenRouterrouter | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio | $3 / 1M input tokensstale | Provider sourceChecked Aug 22, 2026 |
| OpenRouterrouter | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio | $15 / 1M output tokensstale | Provider sourceChecked Aug 22, 2026 |
| Together AIaggregator | Kimi API | model: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio | $3 / 1M input tokensstale | Provider sourceChecked Aug 22, 2026 |
| Together AIaggregator | Kimi API | model: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio | $15 / 1M output tokensstale | Provider sourceChecked Aug 22, 2026 |
Checked evidence, not a ranking: fresh rows were checked against the linked first-party source on the shown date. APIDir does not name a cheapest provider unless exact configurations and units match. Recheck before purchasing.
Developer resources
Provider documentation and pricing links
Compare integration documentation beside the exact first-party source used for each accepted observation.
| Provider | Channel | Developer resources | Accepted evidence |
|---|---|---|---|
| Fireworks AI | aggregator | DocumentationPricing page | 5 observationsEvidence source 1 |
| Kimi API Platform | official | DocumentationPricing page | 3 observationsEvidence source 1 |
| Novita AI | aggregator | DocumentationPricing page | 3 observationsEvidence source 1 |
| OpenRouter | router | DocumentationPricing page | 3 observationsEvidence source 1 |
| Together AI | aggregator | DocumentationPricing page | 2 observationsEvidence source 1 |
Route-specific buying guide
How to budget Kimi K3 on OpenRouter
Route ID: moonshotai/kimi-k3
The accepted rows describe OpenRouter's Kimi K3 route. They are not Moonshot's direct developer prices and do not establish a universal cheapest Kimi provider.
OpenRouter reported a 1,048,576-token context window when checked. Context capacity is a ceiling, not a budget target; retrieval and conversation growth can turn input tokens into the dominant cost.
Cost formula
Estimate one request as uncached input tokens × the input rate, cached input tokens × the cached-input rate, plus output tokens × the output rate. Divide every token count by one million before applying APIDir's displayed rates.
Costs that move the bill
- Long documents, retrieval payloads, conversation history, and agent memory
- Completion length and repeated reasoning, coding, browsing, or tool-use loops
- Cache coverage, cache invalidation frequency, and any route-specific write charge
- Retries, timeouts, fallback models, rate limits, and production support needs
Checks before purchase
- Keep the moonshotai/kimi-k3 route ID and the checked date beside every cost result.
- Benchmark short, median, and long prompts because a single average hides tail cost.
- Cap output and agent steps deliberately, then measure successful-task cost rather than request cost.
- Treat OpenRouter and Moonshot direct access as different offers until their commercial boundaries match.
Source checked Aug 22, 2026: OpenRouter public model API. Recheck the live route before committing spend.
Estimate this route with the LLM workload calculator →
Compare another model family
Decision guide
How to compare Kimi API providers
Kimi is a model family, so a useful provider comparison must name the exact release. The reviewed ranking uses Kimi K3 across OpenRouter, Together AI, and Fireworks AI. Each provider's route spelling remains visible even when the underlying release and token meter match.
Equal token rates do not prove equal service. Context support, cache behavior, throughput, quotas, fallback, endpoint compatibility, regional access, data terms, and support remain separate buying criteria. Use the ranking as a price-meter result and benchmark the service characteristics yourself.
What changes Kimi API cost
- Exact Kimi release, input-output ratio, context, and cache use
- Router or aggregator tier, throughput, rate limits, and any dedicated option
- Tool calls, retries, data retention, regional access, and support
Integration checks before production
- Pin Kimi K3's full provider route ID and reject silent fallback
- Set context, output, timeout, and retry budgets at the caller boundary
- Log token counts, route, terminal status, retry count, and request ID
Workloads this comparison can inform
Long-context assistants
Verify context support and cache meters before extending a simple token estimate.
Batch language processing
Measure real output ratios and concurrency instead of ranking input price alone.
Multi-provider routing
Test fallback identity and error semantics even when the headline rates tie.
How to read the price result
Three reviewed Kimi K3 routes currently publish the same standard input and output rates. APIDir reports the tie for those meters only; it is not a quality, speed, uptime, or support ranking.
Related model price comparisons
Use this evidence safely
A recorded price is not a complete production quote. Open the source, reproduce the exact configuration, and include retries, failed work, storage, egress, support, rate limits, and downstream processing in total cost.
Questions developers ask
What makes two Kimi offers comparable?
The exact model and endpoint, provider route, input and output meters, cache policy, context-length price band, tools, region, and service tier must match.
Does APIDir name a cheapest provider on this page?
Only when at least three explicitly reviewed provider offers are fresh, publication-permitted, and within the same published comparison scope. Stale or mismatched evidence cannot support that claim.
How should I confirm a budget?
Open the linked provider source, verify the configuration, then calculate your own successful and failed workload volume.
APIDir Delta · controlled rollout
Join the source-linked pricing update waitlist.
Confirm your address to join the waitlist. Editorial delivery starts only after sender, suppression, and evidence checks are active.