Moonshot AI

Moonshot AI API Pricing

Kimi series with ultra-long context for document processing

9 paid models · 3 free · Price range: $0.39 - $3.00 /1M

Last updated: Aug 8, 2026

About Moonshot AI

Moonshot AI, a Chinese AI startup, developed the Kimi series known for extremely long context windows (up to 200K+ tokens). Kimi excels at processing long documents, books, and codebases in a single context. Popular in China for document analysis and research assistance.

Key Highlights

  • Ultra-long context windows (200K+)
  • Excellent document processing
  • Strong Chinese language support
  • Competitive pricing
  • Good coding capabilities
Why Choose: Excellent for processing very long documents in a single context. Strong option for Chinese language applications.
12
Total Models
$0.39
Lowest Input
1.0M
Max Context
5
Capabilities

Pricing Features

  • Pay-per-token billing
  • Prompt Caching (discounted)
Pricing Notes:

Competitive pricing for long-context use cases. Good value for document-heavy applications.

API Features

StreamingLong ContextDocument ProcessingCode Generation

Common Use Cases

  • • Long Document Analysis
  • • Book Summarization
  • • Codebase Understanding
  • • Research
  • • Chinese Market

Moonshot AI pricing guide

Focused notes for developers comparing official pricing, API docs, token billing, and model fit.

Moonshot AI pricing for Kimi API models

Moonshot AI pricing is easiest to compare when Kimi model rows, input tokens, output tokens, and context length are kept together. Use the live table below as a model-level planning view, then verify the applicable region, endpoint, and account terms on the official Moonshot or Kimi documentation before production use.

  • Compare input and output token prices separately for long-document workloads.
  • Use context length and document size to explain why a cheaper short-context model may not be the best fit.
  • Check API availability, region, quota, and model-specific billing before committing traffic.
  • Estimate a representative Kimi API workload instead of comparing only the lowest headline rate.

Moonshot AI API pricing versus Kimi subscriptions

Moonshot AI API pricing and Kimi subscription pricing are related but different intents. API usage is estimated from the endpoint and its billable units, while a Kimi consumer or coding plan is a product-access question. This page keeps both topics connected for orientation without treating a subscription as an API allowance.

  • Use this page for API models, token costs, context, and provider-level comparison.
  • Check the official Kimi product pages for consumer, coding, or subscription plan terms.
  • Do not use a monthly plan price as a proxy for token-based API spend.

How to estimate Kimi API cost

Start with monthly input and output tokens, then include long-context prompts, retries, batch behavior, and any multimodal or endpoint-specific charges that apply. For book, codebase, and research workflows, prompt length can dominate the estimate even when the output is short.

  • Measure prompt and completion tokens from a representative document sample.
  • Model repeated context and retry overhead instead of multiplying a single request.
  • Compare Kimi API pricing with other long-context providers using the same token mix.

Related search terms

Moonshot AI API pricing Moonshot AI model pricing Kimi API pricing Kimi AI pricing Kimi subscription pricing

📊 Moonshot AI Model Comparison

Compare all models side by side. Sorted by total price (input + output).

Model Tier Input /1M Output /1M Total /1M Context Best For
Kimi K2 Thinking Budget $0.40 $1.75 $2.15 262k Complex reasoning, math
Kimi K2 0905 (exacto) Budget $0.39 $1.90 $2.29 262k General tasks
Kimi K2.7 Code (batch) Budget $0.47 $2.00 $2.48 262k Complex reasoning, math
Kimi K2.5 Budget $0.45 $2.20 $2.65 262k Complex reasoning, math
Kimi K2 0711 Budget $0.50 $2.40 $2.90 131k General tasks
Kimi K2.7 Code Budget $0.70 $3.50 $4.20 262k Complex reasoning, math

🎯 Which Moonshot AI Model Should You Choose?

Quick recommendations based on your use case.

💰

Lowest Cost

Best value for budget-conscious projects.

Kimi K2 Thinking
$2.15 total
💬

Chat / Customer Service

High volume, short responses.

Kimi K2 Thinking
$2.15 total
💻

Code Generation

Longer outputs, focus on output price.

🧠

Complex Reasoning

Math, logic, multi-step problems.

Kimi K2 Thinking
$2.15 total
👁️

Image Understanding

Analyze images and documents.

📄

Long Documents

Process large files and contexts.

Kimi K3
1.0M context

💰 Moonshot AI Monthly Cost Examples

Estimated monthly costs for common use cases.

Use Case Monthly Usage Kimi K2 Thinking
(Budget)
Kimi K3
(Flagship)
Customer Service Bot
1000 conversations/day
500k input
200k output
$0.55/mo $4.50/mo
Code Assistant
200 requests/day
1.0M input
500k output
$1.27/mo $10.50/mo
Data Analysis
500 analyses/day
2.0M input
300k output
$1.32/mo $10.50/mo

⚔️ Moonshot AI vs Competitors

How does {brand} compare to other major AI providers?

Brand Model Input /1M Output /1M Total /1M Context vs {brand}
Moonshot AI Moonshot AI Kimi K3 Current $3.00 $15.00 $18.00 1.0M
OpenAI OpenAI Codex Mini $1.50 $6.00 $7.50 200k 58% cheaper
OpenAI OpenAI GPT-5.2 Pro $21.00 $168.00 $189.00 400k 950% more
OpenAI OpenAI GPT-5.2 $1.75 $14.00 $15.75 400k 13% cheaper
OpenAI OpenAI GPT-5.1-Codex-Max $1.25 $10.00 $11.25 400k 38% cheaper
OpenAI OpenAI GPT-5.1 $1.25 $10.00 $11.25 400k 38% cheaper
OpenAI OpenAI GPT-5.1-Codex $1.25 $10.00 $11.25 400k 38% cheaper

All Models

Model Input /1M Output /1M Context Capabilities Actions
Kimi K3 $3.00 $15.00 1.0M
chatreasoningtool_usevision
View Details
Kimi K2.7 Code (batch) $0.47 $2.00 262k
chatcodereasoningtool_use
View Details
Kimi K2.7 Code $0.70 $3.50 262k
chatcodereasoningtool_use
View Details
MoonshotAI Kimi Latest $2.50 $14.00 1.0M
chatreasoningtool_usevision
View Details
Kimi K2.6 $0.95 $4.00 262k
chatreasoningtool_usevision
View Details
Kimi K2.5 $0.45 $2.20 262k
chatreasoningtool_usevision
View Details
Kimi K2 Thinking $0.40 $1.75 262k
chatreasoningtool_use
View Details
Kimi K2 0905 (exacto) Cheapest $0.39 $1.90 262k
chattool_use
View Details
Kimi K2 0711 $0.50 $2.40 131k
chattool_use
View Details
Kimi Dev 72B FREE $0.29 $1.15 131k
chatreasoning
View Details
Kimi VL A3B Thinking FREE Free Free 131k
chatvisionreasoning
View Details
Moonlight 16B A3B Instruct FREE Free Free 8k
chat
View Details

❓ Moonshot AI Pricing FAQ

What is the cheapest Moonshot AI model?

The cheapest Moonshot AI model is Kimi K2 Thinking at $2.15 per 1M tokens (input + output combined).

What is the maximum context length for Moonshot AI models?

Moonshot AI models support up to 1.0M context length, allowing you to process large documents and maintain long conversations.

How do I choose between Moonshot AI models?

For budget projects, choose the cheapest model. For code generation, prioritize low output price. For complex reasoning, choose models with reasoning capability. Use our scenario guide above.

Does Moonshot AI support prompt caching?

Yes, Moonshot AI supports prompt caching which can significantly reduce costs for repeated prompts. Check individual model pages for caching prices.

How does Moonshot AI pricing work?

Moonshot AI API pricing is generally evaluated by model, endpoint, input and output usage, context requirements, and account or regional terms. Use the live model rows for a directional estimate and verify the official rate before launch.

What is the difference between Kimi API pricing and Kimi subscription pricing?

Kimi API pricing is usage-based for developer calls, while Kimi subscription pricing covers access to a consumer or coding product. They are separate billing intents and should not be compared as a fixed API token allowance.

How should I estimate Moonshot AI API cost?

Estimate input and output tokens for a representative workload, then account for context length, retries, endpoint rules, region, and any applicable account terms. Long-document and codebase prompts should be measured rather than guessed.