Choosing this model
Selection advice by Y-API. The checks below are suggested evaluations, not published test results.
Where to start
Evaluate a workflow that reconciles a transcript with written notes, keeping evidence attached to each claim. After the text baseline works, test each media format individually before combining them.
What to watch for
An omnimodal model behind a gateway may expose only a subset of its input types. Media tokenization and separate charges can also make the text-token estimate incomplete for audio or video workloads.
Model specifications
These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.
- Context window
- 1,050,000 tokens
- Input
- Text · Audio · Image · Video
- Output
- Text
- Listed on OpenRouter
- Parameters listed by OpenRouter
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Estimate your cost
The current catalog marks this model as free: calls do not deduct credit. This does not promise unlimited capacity, permanent availability or a future free price.
- Estimated credit consumed
- Free — no credit deducted
- Estimated cash equivalent
- Free — no credit deducted
Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.
$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.
| Model | Estimated credit consumed | Estimated cash equivalent |
|---|---|---|
| MiMo-V2.5 | Free — no credit deducted | |
| MiMo-V2.6-Flash | $0.36 | $0.036 |
| Qwen3.8 Flash | $0.45 | $0.045 |
Use the API
These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.
A task to try
Transcript: We agreed to ship Friday. Notes: Ship Thursday. Summarize the conflict, cite both sources and leave the final date unresolved.
Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.
curl --fail-with-body --silent --show-error --max-time 120 \
'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
-H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
-H 'Content-Type: application/json' \
--data-binary @- <<'JSON'
{
"model": "xiaomi/mimo-v2.5",
"messages": [
{
"role": "user",
"content": "Transcript: We agreed to ship Friday. Notes: Ship Thursday. Summarize the conflict, cite both sources and leave the final date unresolved."
}
],
"max_tokens": 4096
}
JSONHow to evaluate the result
The answer should preserve the disagreement instead of inventing a resolution. For audio tests, use a recording with a known transcript and check numbers, names and negations.
Before you choose
Can MiMo-V2.5 generate audio from this API?
The reviewed catalog lists text output, not audio output. Audio as an input modality does not establish speech synthesis, and Y-API media compatibility requires a separate check.
Sources & scope
Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.