Kimi K3

moonshotai/kimi-k3

Kimi K3 is Moonshot AI’s open-weight multimodal model for coding, knowledge work and long-horizon agents. Its published design emphasizes iterating against repositories, images, logs and runtime feedback rather than generating a single isolated answer.

Input · Account credit
$3.00

per 1M tokens

Cash equivalent $0.30

Output · Account credit
$15.00

per 1M tokens

Cash equivalent $1.50

Context window
1,048,576

tokens · OpenRouter

Input → Output
Text · Image · Video→ Text
Model specifications
Sources reviewed: Y-API catalog snapshot:

Choosing this model

Selection advice by Y-API. The checks below are suggested evaluations, not published test results.

Where to start

Evaluate it in a repository task with a real feedback loop: locate a failure, propose a patch, run tests and revise. Score completed requirements and unintended changes, not the length of the plan.

What to watch for

An agent harness supplies tools, state and execution limits; the model does not acquire those through this API call alone. Its larger context also makes repeated full-history requests worth budgeting explicitly.

Model specifications

These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.

Context window
1,048,576 tokens
Input
Text · Image · Video
Output
Text
Listed on OpenRouter
Parameters listed by OpenRouter
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter model page & specifications

Estimate your cost

Estimated credit consumed
$10.50
Estimated cash equivalent
$1.05

Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.

$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.

Same token budget, different models
ModelEstimated credit consumedEstimated cash equivalent
Kimi K3$10.50$1.05
Kimi K2.6$2.95$0.295
Claude Opus 5$17.50$1.75

Use the API

These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.

A task to try

An API times out only on large exports. Design a diagnostic sequence that distinguishes slow SQL, serialization cost and proxy timeouts, without changing production data.

Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.

cURL · Kimi K3
curl --fail-with-body --silent --show-error --max-time 120 \
  'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
  -H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
  -H 'Content-Type: application/json' \
  --data-binary @- <<'JSON'
{
  "model": "moonshotai/kimi-k3",
  "messages": [
    {
      "role": "user",
      "content": "An API times out only on large exports. Design a diagnostic sequence that distinguishes slow SQL, serialization cost and proxy timeouts, without changing production data."
    }
  ],
  "max_tokens": 4096
}
JSON

How to evaluate the result

Require measurements that isolate the three stages and a safe order of investigation. In an agent run, check that each tool result changes the next decision rather than being ignored.

Before you choose

Does calling K3 automatically start multiple agents?

No. Multi-agent behavior requires orchestration in your application or client. Set tool permissions, budgets and stopping rules separately from the model selection.

Sources & scope

Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.