Moonshot AI · text · family comparison
Kimi API pricing by model and provider
Compare Kimi API pricing across provider routes without mixing Kimi versions, token meters, context tiers, or service channels.
Model-level provider comparison
Kimi API price ranking status
Current answer: APIDir currently publishes 2 comparable price rankings. Keep each meter and configuration separate before choosing a route.
- Provider routes with evidence
- 3
- Accepted observations
- 7
- Billing meters
- 1M cached input tokens, 1M input tokens, 1M output tokens
- Latest accepted check
- Aug 22, 2026
Kimi K3 API input token pricing
Kimi K3 input token price ranking
- Reviewed comparison scope
- Kimi K3 · shared API route · global · USD per 1M input tokens
- Publication gate
- At least 3 unique providers with accepted, fresh, reviewed evidence
| Price rank | Provider | Route and tier | Published meter | Evidence |
|---|---|---|---|---|
| 1Tie | Fireworks AI | aggregator · serverless-standard | $3 / 1M input tokensfresh | First-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026 |
| 1Tie | OpenRouter | router · Kimi K3 standard | $3 / 1M input tokensfresh | First-party sourceChecked Aug 22, 2026Fresh until Aug 25, 2026 |
| 1Tie | Together AI | aggregator · serverless | $3 / 1M input tokensfresh | First-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026 |
Price meter only: The three providers identify Kimi K3 and the input token meter. Tier names and service features remain visible and are not treated as equivalent quality, speed, or reliability. Equal prices share a price rank; provider name and offer ID determine only their stable display order.
Kimi K3 API output token pricing
Kimi K3 output token price ranking
- Reviewed comparison scope
- Kimi K3 · shared API route · global · USD per 1M output tokens
- Publication gate
- At least 3 unique providers with accepted, fresh, reviewed evidence
| Price rank | Provider | Route and tier | Published meter | Evidence |
|---|---|---|---|---|
| 1Tie | Fireworks AI | aggregator · serverless-standard | $15 / 1M output tokensfresh | First-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026 |
| 1Tie | OpenRouter | router · Kimi K3 standard | $15 / 1M output tokensfresh | First-party sourceChecked Aug 22, 2026Fresh until Aug 25, 2026 |
| 1Tie | Together AI | aggregator · serverless | $15 / 1M output tokensfresh | First-party sourceChecked Aug 22, 2026Fresh until Aug 29, 2026 |
Price meter only: The order covers the listed output token meter only. Context, cache, throughput, quotas, endpoint behavior, and support remain separate buying criteria. Equal prices share a price rank; provider name and offer ID determine only their stable display order.
Uncollapsed source records
All accepted offer evidence
These rows remain separate until their offered model ID, route, meter, tier, resolution, duration, audio, and region are proven comparable.
| Provider | Model | Configuration | Recorded price | Evidence |
|---|---|---|---|---|
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio | $3 / 1M input tokensfresh | Provider sourceChecked Aug 22, 2026 |
| Fireworks AIaggregator | Kimi API | model: fireworks/kimi-k3 · chat-completions · serverless-standard · not-applicable · no audio | $15 / 1M output tokensfresh | Provider sourceChecked Aug 22, 2026 |
| OpenRouterrouter | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio | $0.3 / 1M cached input tokensfresh | Provider sourceChecked Aug 22, 2026 |
| OpenRouterrouter | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio | $3 / 1M input tokensfresh | Provider sourceChecked Aug 22, 2026 |
| OpenRouterrouter | Kimi API | model: moonshotai/kimi-k3 · chat-completions · Kimi K3 standard · not-applicable · no audio | $15 / 1M output tokensfresh | Provider sourceChecked Aug 22, 2026 |
| Together AIaggregator | Kimi API | model: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio | $3 / 1M input tokensfresh | Provider sourceChecked Aug 22, 2026 |
| Together AIaggregator | Kimi API | model: moonshotai/Kimi-K3 · chat-completions · serverless · not-applicable · no audio | $15 / 1M output tokensfresh | Provider sourceChecked Aug 22, 2026 |
Checked evidence, not a ranking: fresh rows were checked against the linked first-party source on the shown date. APIDir does not name a cheapest provider unless exact configurations and units match. Recheck before purchasing.
Developer resources
Provider documentation and pricing links
Compare integration documentation beside the exact first-party source used for each accepted observation.
| Provider | Channel | Developer resources | Accepted evidence |
|---|---|---|---|
| Fireworks AI | aggregator | DocumentationPricing page | 2 observationsEvidence source 1 |
| OpenRouter | router | DocumentationPricing page | 3 observationsEvidence source 1 |
| Together AI | aggregator | DocumentationPricing page | 2 observationsEvidence source 1 |
Route-specific buying guide
How to budget Kimi K3 on OpenRouter
Route ID: moonshotai/kimi-k3
The accepted rows describe OpenRouter's Kimi K3 route. They are not Moonshot's direct developer prices and do not establish a universal cheapest Kimi provider.
OpenRouter reported a 1,048,576-token context window when checked. Context capacity is a ceiling, not a budget target; retrieval and conversation growth can turn input tokens into the dominant cost.
Cost formula
Estimate one request as uncached input tokens × the input rate, cached input tokens × the cached-input rate, plus output tokens × the output rate. Divide every token count by one million before applying APIDir's displayed rates.
Costs that move the bill
- Long documents, retrieval payloads, conversation history, and agent memory
- Completion length and repeated reasoning, coding, browsing, or tool-use loops
- Cache coverage, cache invalidation frequency, and any route-specific write charge
- Retries, timeouts, fallback models, rate limits, and production support needs
Checks before purchase
- Keep the moonshotai/kimi-k3 route ID and the checked date beside every cost result.
- Benchmark short, median, and long prompts because a single average hides tail cost.
- Cap output and agent steps deliberately, then measure successful-task cost rather than request cost.
- Treat OpenRouter and Moonshot direct access as different offers until their commercial boundaries match.
Source checked Aug 22, 2026: OpenRouter public model API. Recheck the live route before committing spend.
Estimate this route with the LLM workload calculator →
Compare another model family
Decision guide
How to compare Kimi API API providers
Kimi is a model family, so a useful provider comparison must name the exact release. The reviewed ranking uses Kimi K3 across OpenRouter, Together AI, and Fireworks AI. Each provider's route spelling remains visible even when the underlying release and token meter match.
Equal token rates do not prove equal service. Context support, cache behavior, throughput, quotas, fallback, endpoint compatibility, regional access, data terms, and support remain separate buying criteria. Use the ranking as a price-meter result and benchmark the service characteristics yourself.
What changes the Kimi API API cost
- Exact Kimi release, input-output ratio, context, and cache use
- Router or aggregator tier, throughput, rate limits, and any dedicated option
- Tool calls, retries, data retention, regional access, and support
Integration checks before production
- Pin Kimi K3's full provider route ID and reject silent fallback
- Set context, output, timeout, and retry budgets at the caller boundary
- Log token counts, route, terminal status, retry count, and request ID
Workloads this comparison can inform
Long-context assistants
Verify context support and cache meters before extending a simple token estimate.
Batch language processing
Measure real output ratios and concurrency instead of ranking input price alone.
Multi-provider routing
Test fallback identity and error semantics even when the headline rates tie.
How to read the price result
Three reviewed Kimi K3 routes currently publish the same standard input and output rates. APIDir reports the tie for those meters only; it is not a quality, speed, uptime, or support ranking.
Related model price comparisons
Use this evidence safely
A recorded price is not a complete production quote. Open the source, reproduce the exact configuration, and include retries, failed work, storage, egress, support, rate limits, and downstream processing in total cost.
Questions developers ask
What makes two Kimi offers comparable?
The exact model and endpoint, provider route, input and output meters, cache policy, context-length price band, tools, region, and service tier must match.
Does APIDir name a cheapest provider on this page?
Only when at least three explicitly reviewed provider offers are fresh, publication-permitted, and within the same published comparison scope. Stale or mismatched evidence cannot support that claim.
How should I confirm a budget?
Open the linked provider source, verify the configuration, then calculate your own successful and failed workload volume.
APIDir Delta · controlled rollout
Reserve source-linked price research.
Confirm your address to join the rollout. Editorial delivery starts only after sender, suppression, and evidence checks are active.