MiniMax M3

minimax/minimax-m3

MiniMax M3 is the multimodal successor to M2.7: text, image and video input, text output, and a 1M-token window aimed at long-horizon agentic and coding work. Unlike the free M2.7 entry it is billed per token.

Input · Account credit
$0.30

per 1M tokens

Cash equivalent $0.03

Output · Account credit
$1.20

per 1M tokens

Cash equivalent $0.12

Context window
1,048,576

tokens · OpenRouter

Input → Output
Text · Image · Video→ Text
Model specifications
Sources reviewed: Y-API catalog snapshot:

Choosing this model

Selection advice by Y-API. The checks below are suggested evaluations, not published test results.

Where to start

Try it on agent loops that have to run for many steps without losing the thread — a build-fix-retest cycle, or a data pipeline that must keep its invariants across transformations.

What to watch for

The listed context window is the model’s, not a promise about what the gateway will accept; long inputs also get expensive fast. Video input in particular needs its own compatibility test before you build on it.

Model specifications

These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.

Context window
1,048,576 tokens
Input
Text · Image · Video
Output
Text
Listed on OpenRouter
Parameters listed by OpenRouter
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter model page & specifications

API pricing & cost estimate

Estimated credit consumed
$0.90
Estimated cash equivalent
$0.09

Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.

$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.

Same token budget, different models
ModelEstimated credit consumedEstimated cash equivalent
MiniMax M3$0.90$0.09
MiniMax M2.7Free — no credit deducted
Qwen3.8 Flash$0.45$0.045
View LLM API pricing & billing

API integration examples

These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.

A task to try

A migration script must be safe to re-run. List the operations that would duplicate data on a second run and rewrite each one to be idempotent.

Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.

cURL · MiniMax M3
curl --fail-with-body --silent --show-error --max-time 120 \
  'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
  -H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
  -H 'Content-Type: application/json' \
  --data-binary @- <<'JSON'
{
  "model": "minimax/minimax-m3",
  "messages": [
    {
      "role": "user",
      "content": "A migration script must be safe to re-run. List the operations that would duplicate data on a second run and rewrite each one to be idempotent."
    }
  ],
  "max_tokens": 4096
}
JSON

How to evaluate the result

Run the rewritten script twice against the same fixture and diff the results. Any difference between the two runs means an operation was not actually made idempotent.

Before you choose

Is M3 simply a bigger M2.7?

No. M3 adds image and video input and a much longer context window, and it is a paid entry where M2.7 is free. Treat it as a different model to evaluate, not a drop-in upgrade.

Sources & scope

Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.