Kimi K3 is here: our most capable model

Meet Kimi K3, our most capable model.

Kimi K3 is a 2.8-trillion-parameter model with native vision and a 1M-token context window, built for long-horizon coding, knowledge work, and reasoning. It can sustain long engineering sessions with minimal oversight, navigate massive repositories, and take on work that used to be out of reach.

  • Model ID: kimi-k3

  • Pricing: $0.30/MTok cached input · $3.00/MTok uncached input · $15.00/MTok output — 90%+ cache hit rate in coding, so most input costs $0.30.

  • Get started: Kimi K3 - Kimi API Platform

Use the top-level reasoning_effort field to choose how hard K3 thinks: low for the fastest responses, high for a balance of speed and quality, max for the hardest tasks.

Two quick notes: pass the full assistant message back in multi-turn conversations, and don’t switch to K3 mid-session from another model. K3 is also proactive by nature — if you need tighter boundaries, set them explicitly in your system prompt or AGENTS.md.

1 Like