Kimi K2
moonshot/kimi-k2Moonshot's open agentic flagship, long-context stamina and tool use that holds up across extended sessions.
Model overview
Context window
256K
tokens
Input price
—
quoted on request
Output price
—
quoted on request
Weekly volume
—
tokens / week
Open weights
Yes
self-hostable
Capabilities
5
of 11 tags
Description
Kimi K2 is a language model from Moonshot AI. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. Moonshot AI’s stated focus is long-context agentic intelligence (Kimi). Its mean it can also be self-hosted under its licence.
Capabilities
- Tool calling
- Calls functions you define and returns their arguments as structured data, the basis of agents.
- JSON mode
- Constrains the response to valid JSON matching a schema you supply.
- Streaming
- Emits tokens as they are generated, so answers appear progressively rather than all at once.
- Open weights
- The parameters are published, so the model can be inspected, fine-tuned and self-hosted under its licence.
- Caching
- Reuses already-processed prompt prefixes across requests, cutting input cost for repeated context.
Strengths & weaknesses
Strengths
- Long-context quality that survives real use
- Agentic tool use
- Open weights
Weaknesses
- Output price above Chinese peers
- Vision requires the separate VL model
Pricing
Free| Rate | Price | Unit |
|---|---|---|
| Input | $0.00 | per 1M tokens |
| Output | — not charged | — |
Prompt caching is supported, so repeated prefixes bill below the listed input rate. Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.
Context window
256Ktokens
covers prompt and response together, so a long input leaves less room for the answer.
Supported features
| Feature | Support |
|---|---|
| Tool calling | Supported |
| JSON mode | Supported |
| Streaming | Supported |
| Vision input | Not supported |
| Audio input | Not supported |
| Long context | Not supported |
| Open weights | Supported |
| Reasoning | Not supported |
| Fine-tunable | Not supported |
| Batch | Not supported |
| Caching | Supported |
Example use cases
Long-document analysis
Long-context quality that survives real use
Agent workflows
Agentic tool use
Research assistants
Open weights
Comparison
| Attribute | Kimi K2 | Kimi K2 Thinking | DeepSeek V3.2 |
|---|---|---|---|
| Context window | 256K | 256K | 128K |
| Input price | $0.00 | $0.00 | $0.00 |
| Output price | — not charged | — not charged | — not charged |
| Tool calling | Supported | Supported | Supported |
| JSON mode | Supported | Not supported | Supported |
| Vision input | Not supported | Not supported | Not supported |
| Open weights | Supported | Supported | Supported |
| Prompt caching | Supported | Not supported | Supported |
Best value in each row is highlighted. Illustrative placeholder data.
Provider information
View Moonshot AI →Moonshot AI · 月之暗面 · Beijing, China · est. 2023
Moonshot's Kimi line made long context its signature, and its K2 generation paired that with serious agentic tool use, released with open weights. The lab moves fast and publishes aggressively, making Kimi one of the most-watched model families in the ecosystem.
Recent releases
Documentation
Frequently asked questions
Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Prompt caching bills repeated prefixes below the listed input rate. Every figure on this page is an illustrative placeholder.