chat.kimi_k2_instruct
moonshotai/kimi-k2-instruct

The Kimi K2 Instruct API is Moonshot AI's standard instruction-tuned, open-weight Kimi K2 model — fast direct answers, reliable instruction-following, and native tool use, with structured outputs and a long context window. The Kimi K2 Instruct API is OpenAI-compatible: point any existing SDK at RouterBase and set the model to kimi-k2-instruct. Input at $0.48 / 1M tokens, output at $1.95 / 1M tokens — 15% below the official rate, all through one REST endpoint.

Input
Kimi K2 Instruct
max_tokens4096
temperature1
top_p1
presence_penalty0
frequency_penalty0
Kimi K2 Instruct Online
Hi! I'm a helpful AI assistant. What can I do for you?

Kimi K2 Instruct API: Chat

Use the Kimi K2 Instruct API to run Moonshot AI's standard instruction-tuned Kimi K2 — fast, direct answers with reliable instruction-following and native tool use, no extended thinking, at open-model prices.

The Kimi K2 Instruct API is Moonshot AI's standard instruction-tuned, open-weight Kimi K2 — fast, direct answers with reliable instruction-following and native tool use, plus structured outputs and a long context window. Routed through RouterBase, the Kimi K2 Instruct API is fully OpenAI-compatible: point any existing SDK at RouterBase's base URL, set the model to kimi-k2-instruct, and your first Kimi K2 Instruct API call goes through immediately.

Pricing for the Kimi K2 Instruct API is $0.48 / 1M input and $1.95 / 1M output — 15% below the official rate, via one RouterBase key (no Moonshot AI account).

Why this model

Six reasons teams ship on the Kimi K2 Instruct API

From snappy, direct answers to a 15%-off price — what makes the Kimi K2 Instruct API the everyday Kimi K2.

Direct, low-latency answers

Kimi K2 Instruct replies straight away — no hidden reasoning tokens to pay for or wait on, which makes it the fast default for chat, tools, and high-throughput apps.

Follows instructions

Give the Kimi K2 Instruct API a format, a system prompt, or a tool schema and it sticks to it — the instruction-tuned Kimi K2 built for predictable, do-what-I-said output.

Native tool calling

The Kimi K2 Instruct API does function / tool calling in the OpenAI format — Kimi K2 was built around tool use, so agents drop in without custom glue.

Long context window

The Kimi K2 Instruct API takes a long context — feed it whole documents, codebases, or full conversation history in one call.

OpenAI-compatible, one key

The Kimi K2 Instruct API uses the OpenAI chat-completions format, so one RouterBase key swaps Kimi K2 Instruct in beside GPT-5.5, Claude, Gemini, and 200+ other models.

15% off the list price

Kimi K2 Instruct runs 15% under Moonshot AI's published rate — $0.48 / 1M input, $1.95 / 1M output. Open-weight pricing, billed per token, nothing per request.

RouterBase dashboard preview
Quickstart

Get started with the Kimi K2 Instruct API in 3 steps

From sign-up to your first response in under 5 minutes.

  1. Create a RouterBase API key

    Sign up and create an API key — it reaches Kimi K2 Instruct and every other model in the RouterBase catalog from one account.

  2. Send your first message

    Point any OpenAI-compatible SDK at routerbase.com/v1 and set the model to kimi-k2-instruct. Streaming, tool calls, and structured outputs all follow the standard chat-completions schema.

  3. Inspect usage

    Every Kimi K2 Instruct API response returns a token breakdown — input and output — so cost stays visible on every call.

Pricing

Pay only for what you use

RouterBase passes through partner-tier pricing. Compared against the model's official published API rate.

Customer stories

What teams build with the Kimi K2 Instruct API

Real production loads running on the RouterBase model catalog across 200+ models.

Marcus Reyes
Marcus ReyesCTO, Paradigm AI

Kimi K2 Instruct is our default for high-throughput chat — direct answers, no thinking tokens to pay for, and the bill stays low.

Priya Lakshmi
Priya LakshmiFounder, Quillo

We route most user requests to the Kimi K2 Instruct API — it follows our format instructions every time and responds fast.

Aoi Tanaka
Aoi TanakaML Lead, Daybreak Robotics

Tool calls on the Kimi K2 Instruct API drop straight into our agent loop — instruction-tuned, predictable, same OpenAI format.

Ethan Nguyen
Ethan NguyenHead of Engineering, Compound Studio

We A/B the Kimi K2 Instruct API against GPT and Claude behind one RouterBase key — for direct tasks it matches them at a fraction of the cost.

Lucas Fernandes
Lucas FernandesEngineering Manager, Light

Pointing our OpenAI SDK at the Kimi K2 Instruct API took one base-URL change, and latency dropped versus the thinking models we tried.

Thomas Beck
Thomas BeckStaff Engineer, Northbeam

Structured outputs from the Kimi K2 Instruct API come back clean — it does what the schema says without a reasoning detour.

Jonas Keller
Jonas KellerIndie Developer

As a solo dev the Kimi K2 Instruct API is my workhorse — fast, cheap, and it follows instructions without surprises.

Sophia Martín
Sophia MartínCTO, Relay

For our high-volume endpoints the Kimi K2 Instruct API is the best cost-per-answer we tested — quick and on-instruction.

David Okonkwo
David OkonkwoCo-founder, Figment

Per-token billing plus no thinking tokens means the Kimi K2 Instruct API keeps our cost flat at scale.

Frequently Asked Questions

Common questions about the Kimi K2 Instruct API.

The Kimi K2 Instruct API is RouterBase's pass-through to Moonshot AI's Kimi K2 Instruct — the standard instruction-tuned, open-weight Kimi K2 for fast direct answers, instruction-following, and tool use, served over an OpenAI-compatible REST interface with native tool calling and structured outputs.