MiMo-V2.5

xiaomi/mimo-v2.5Free — no credit deducted

Xiaomi MiMo-V2.5 is an omnimodal model whose reference entry lists text, image, audio and video inputs. Its output is text: broad perception support should not be confused with speech or image generation.

Input · Account credit
Free — no credit deducted

per 1M tokens

Output · Account credit
Free — no credit deducted

per 1M tokens

Context window
1,050,000

tokens · OpenRouter

Input → Output
Text · Audio · Image · Video→ Text
Model specifications
Sources reviewed: Y-API catalog snapshot:

Choosing this model

Selection advice by Y-API. The checks below are suggested evaluations, not published test results.

Where to start

Evaluate a workflow that reconciles a transcript with written notes, keeping evidence attached to each claim. After the text baseline works, test each media format individually before combining them.

What to watch for

An omnimodal model behind a gateway may expose only a subset of its input types. Media tokenization and separate charges can also make the text-token estimate incomplete for audio or video workloads.

Model specifications

These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.

Context window
1,050,000 tokens
Input
Text · Audio · Image · Video
Output
Text
Listed on OpenRouter
Parameters listed by OpenRouter
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
OpenRouter model page & specifications

Estimate your cost

The current catalog marks this model as free: calls do not deduct credit. This does not promise unlimited capacity, permanent availability or a future free price.

Estimated credit consumed
Free — no credit deducted
Estimated cash equivalent
Free — no credit deducted

Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.

$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.

Same token budget, different models
ModelEstimated credit consumedEstimated cash equivalent
MiMo-V2.5Free — no credit deducted
MiMo-V2.6-Flash$0.36$0.036
Qwen3.8 Flash$0.45$0.045

Use the API

These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.

A task to try

Transcript: We agreed to ship Friday. Notes: Ship Thursday. Summarize the conflict, cite both sources and leave the final date unresolved.

Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.

cURL · MiMo-V2.5
curl --fail-with-body --silent --show-error --max-time 120 \
  'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
  -H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
  -H 'Content-Type: application/json' \
  --data-binary @- <<'JSON'
{
  "model": "xiaomi/mimo-v2.5",
  "messages": [
    {
      "role": "user",
      "content": "Transcript: We agreed to ship Friday. Notes: Ship Thursday. Summarize the conflict, cite both sources and leave the final date unresolved."
    }
  ],
  "max_tokens": 4096
}
JSON

How to evaluate the result

The answer should preserve the disagreement instead of inventing a resolution. For audio tests, use a recording with a known transcript and check numbers, names and negations.

Before you choose

Can MiMo-V2.5 generate audio from this API?

The reviewed catalog lists text output, not audio output. Audio as an input modality does not establish speech synthesis, and Y-API media compatibility requires a separate check.

Sources & scope

Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.