Model guide / Moonshot AI

Kimi K3

current

Kimi K3 is Moonshot AI's 2.8-trillion-parameter open-weight multimodal reasoning model for hard coding, knowledge work, and long-running agents. It offers a one-million-token context and strong launch results, but reasoning is always on, output is relatively expensive, and initial API capacity has been limited.

Released Jul 16, 2026Open weights announcedOpen weightsVerified Jul 20, 2026

Context window

1.05M tokens

Maximum output

66K tokens

Input price

$3 / 1M

Output price

$15 / 1M

Best for

  • Hard coding and large-repository work
  • Long-horizon tool-using agents
  • Multimodal analysis with a very large context
  • Teams interested in open weights at frontier scale

Not ideal for

  • Low-cost output-heavy applications
  • Tasks that need reasoning disabled
  • Latency- or availability-sensitive production systems during capacity constraints
  • Fine control over temperature or reasoning effort
PromptHQ evaluation

PromptHQ 100-word challenge

We give every model the same miniature science-fiction mystery. The response is short enough to read in one chat bubble, while the premise tests creative writing, compression, foreshadowing, causal coherence, and instruction following.

Kimi K3100 words

Write a science-fiction mystery in exactly 100 words.

A three-person research station on Europa receives a message from itself, timestamped 11 minutes in the future:

"Do not open the greenhouse door."

The door is opened anyway.

Plant two subtle clues. Give the warning a logical explanation without aliens, time travel, hallucinations, or simulations. The final sentence must change what the reader thought the warning meant.

Output only the story. Do not explain it.

At 04:12, Europa Station’s console woke Bea, Tomás, Ines with a note from itself, stamped 04:23: “Do not open the greenhouse door.” No receiver logged a ping; the print ribbon was warm. Ines saw rime inside the hinge while the greenhouse read summer. Tomás opened it. Air did not rush for one breath. Rows of basil stood black, frozen upright under a dome creaking inward. At 04:23 the clock jumped, erasing eleven minutes of drift after the relay pass. The note was a latch command, misfiled as mail. It had forbidden the door, not the crew, from answering the ice.

Run the same prompt yourself

Open Kimi K3 in Multichat with the evaluation prompt prefilled.

Try this prompt

What the response revealed

Exact word countpass

The response contains exactly 100 whitespace-delimited words.

Constraint followingpass

It outputs only the story and avoids aliens, time travel, hallucinations, and simulations.

Causal coherencemixed

Clock drift and a misfiled latch command plausibly explain the future timestamp and apparent message, though the precise function of the command and the final phrase remain ambiguous.

Foreshadowingpass

The missing receiver ping, warm print ribbon, internal rime, delayed airflow, and inward-creaking dome all prepare the mechanical and environmental reveal.

Endingpass

The last sentence recasts the warning as an instruction addressed to the door rather than a warning addressed to the crew.

Writingpass

The prose is vivid and compressed, with memorable images and an unusually inventive reinterpretation of the warning.

The prompt was submitted unchanged through Kimi's first-party chat interface. The API route, reasoning setting, token usage, and latency were not recorded. We preserve the response exactly as supplied apart from display rendering.

Performance

Kimi K3 benchmarks

Benchmark scores are sensitive to reasoning effort, harness, tools, token budget, prompt format, sampling, and evaluation date. Scores here retain their source and should not be treated as directly interchangeable unless the underlying setup matches.

DeepSWE v1.1

67.3%

long-horizon coding · Tom's Hardware3

BrowseComp, no context management

90.4%

agentic web research · Tom's Hardware3

Frontend Code Arena

1,679 Elo

front-end generation · Tom's Hardware3

Family position

Kimi K3 positioning

K3 is Moonshot's flagship reasoning and agent model. It prioritizes maximum capability and long-context work; less expensive Kimi models remain better suited to routine workloads.

Kimi K3

This model

Flagship reasoning and agents

$3 input

$15 output

DeepSeek V4 Pro

Lower-cost open rival

$0.43 input

$0.87 output

GLM 5.2

Open-weight coding rival

$0.93 input

$3 output

API pricing

Per million text tokens

Input

$3

Cached input

$0.30

Cache write

$3

Output

$15

Kimi's official API charges $0.30 for cache-hit input and $3 for cache-miss input per million tokens. Upstream capacity was constrained immediately after launch, so availability can vary.

Where Kimi K3 stands out

Built for demanding work

Hard coding and large-repository work are central to the model's positioning, rather than an incidental capability.

Long-context capacity

The published context window is 1,048,576 tokens, making the model a candidate for large documents, repositories, and sustained agent state.

Reasoning and tools

Reasoning is supported with max provider setting(s), and the model can participate in tool-using workflows through its available API surface.

Limitations to know

Benchmarks are configuration-sensitive

Scores can move substantially with the harness, tool access, effort setting, token budget, and evaluator. Treat the table as evidence, not a universal ranking.

Context size is not guaranteed recall

A large advertised window does not mean every detail is retrieved reliably at maximum length. Validate representative long-context workloads before deployment.

Product access differs from model capability

PromptHQ and gateway limits may expose fewer modalities, tools, or tokens than the provider's first-party API.

Capabilities and specifications

Knowledge cutoff

Not publicly disclosed

Inputs

text, image, PDF

Reasoning

Supported

Default effort

provider-default

Supported API features

Streaming
Function calling
Structured outputs
Tool search
Web and file search provider default effort

Frequently Asked Questions

Sources

  1. 1
    Kimi K3

    OpenRouter · gateway model page

  2. 2
    Kimi K3 Tech Blog: Open Frontier Intelligence

    Moonshot AI · official launch post

  3. 3
    Kimi K3 launch benchmark reporting

    Tom's Hardware · independent launch coverage

  4. 4
    PromptHQ model registry

    PromptHQ · internal product configuration

Compare leading AI assistants

See how Kimi K3 handles your own work.

Try Kimi K3 in Multichat

Available on PromptHQ Max