# MiMo-V2.6-Pro — API pricing & integration guide | Y-API

> MiMo-V2.6-Pro is Xiaomi’s flagship open-weight entry, described by the publisher as a model above one trillion parameters. It accepts text, image, video and audio input over a 1,050,000-token window and returns text.

This is the markdown representation of https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.6-pro. Generated by `scripts/generate-seo-assets.mjs` from the same copy the page renders — do not edit by hand.

Model ID: `xiaomi/mimo-v2.6-pro`

Sources reviewed: 2026-10-11. Y-API catalog snapshot: 2026-10-11.

## Choosing this model

Selection advice by Y-API. The checks below are suggested evaluations, not published test results.

### Where to start

Evaluate it on demanding multimodal work where the Flash variant has fallen short: video understanding, audio-plus-text reasoning, or tasks that need the larger parameter count to hold a complex specification.

### What to watch for

A published parameter count describes the checkpoint, not the throughput or quality you will see through the gateway. Audio and video input still need their own compatibility test before you build on them.

## Model specifications

These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.

| Field | Value |
| --- | --- |
| Context window | 1,050,000 tokens |
| Input | Text, Image, Video, Audio |
| Output | Text |
| Listed on OpenRouter | 2026-09-21 |
| Parameters listed by OpenRouter | `frequency_penalty`, `include_reasoning`, `logit_bias`, `max_tokens`, `min_p`, `presence_penalty`, `reasoning`, `repetition_penalty`, `response_format`, `seed`, `stop`, `structured_outputs`, `temperature`, `tool_choice`, `tools`, `top_k`, `top_p` |

## API pricing & cost estimate

| Price basis | Input / 1M tokens | Output / 1M tokens |
| --- | --- | --- |
| Account credit | $0.50 | $1.00 |
| Cash equivalent | $0.05 | $0.10 |

$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.

Input tokens / request: 1000. Output tokens / request: 500. Number of requests: 1000.

Estimated credit consumed: $1.00. Estimated cash equivalent: $0.10.

Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.

### Same token budget, different models

| Model | Estimated credit consumed | Estimated cash equivalent |
| --- | --- | --- |
| [MiMo-V2.6-Pro](https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.6-pro) | $1.00 | $0.10 |
| [MiMo-V2.6-Flash](https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.6-flash) | $0.36 | $0.036 |
| [MiMo-V2.5](https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.5) | Free — no credit deducted | Free — no credit deducted |

[View LLM API pricing & billing](https://y-api.bestvirtualgoods.com/pricing)

## API integration examples

These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.

Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.

### A task to try

Given a screen recording and its narration, list the steps the user actually performed, in order, and flag any step the narration does not mention.

```bash
curl --fail-with-body --silent --show-error --max-time 120 \
  'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
  -H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
  -H 'Content-Type: application/json' \
  --data-binary @- <<'JSON'
{
  "model": "xiaomi/mimo-v2.6-pro",
  "messages": [
    {
      "role": "user",
      "content": "Given a screen recording and its narration, list the steps the user actually performed, in order, and flag any step the narration does not mention."
    }
  ],
  "max_tokens": 4096
}
JSON
```

```python
import json
import os
import sys
import urllib.error
import urllib.request

payload = {
  "model": "xiaomi/mimo-v2.6-pro",
  "messages": [
    {
      "role": "user",
      "content": "Given a screen recording and its narration, list the steps the user actually performed, in order, and flag any step the narration does not mention."
    }
  ],
  "max_tokens": 4096
}

request = urllib.request.Request(
    "https://api.y-api.bestvirtualgoods.com/v1/chat/completions",
    data=json.dumps(payload).encode("utf-8"),
    headers={
        "Authorization": "Bearer " + os.environ["TOKEN"],
        "Content-Type": "application/json",
    },
    method="POST",
)
try:
    with urllib.request.urlopen(request, timeout=120) as response:
        result = json.load(response)
    print(json.dumps(result, ensure_ascii=False, indent=2))
except urllib.error.HTTPError as error:
    print(error.read().decode("utf-8"), file=sys.stderr)
    raise SystemExit(1)
```

### How to evaluate the result

Compare the extracted steps against a manually written ground truth. Then run the same recording through the Flash entry to see whether the Pro tier is worth its higher price for your footage.

## Before you choose

### Does the linked card describe the served weights?

The card names a checkpoint from the publisher’s training pipeline. It is useful technical background, but it does not prove which weights the gateway serves, so verify behavior with your own requests.

## Sources & scope

Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.

- [OpenRouter model page & specifications](https://openrouter.ai/xiaomi/mimo-v2.6-pro)
- [Publisher model card linked by OpenRouter](https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Pro-RL)
- [Y-API catalog & prices (JSON)](https://y-api.bestvirtualgoods.com/models.json)

## Models to compare

Compare these alternatives on the same inputs. Their descriptions explain different roles; a lower price does not establish equivalent quality.

### [MiMo-V2.6-Flash](https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.6-flash)

MiMo-V2.6-Flash is Xiaomi’s efficiency-oriented multimodal entry for coding and agent workflows. OpenRouter links the Flash-RL checkpoint card, which describes reinforcement-learning work across varied task environments.

### [MiMo-V2.5](https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.5)

Xiaomi MiMo-V2.5 is an omnimodal model whose reference entry lists text, image, audio and video inputs. Its output is text: broad perception support should not be confused with speech or image generation.

## Links

- HTML version of this page: https://y-api.bestvirtualgoods.com/models/xiaomi/mimo-v2.6-pro
- Site index for agents: https://y-api.bestvirtualgoods.com/llms.txt
- Full reference (single file): https://y-api.bestvirtualgoods.com/llms-full.txt
- OpenAPI 3.1 spec: https://y-api.bestvirtualgoods.com/openapi.json
- Model catalog (JSON, no key needed): https://y-api.bestvirtualgoods.com/models.json
- API base URL: `https://api.y-api.bestvirtualgoods.com/v1`
- Contact: support@bestvirtualgoods.com
- Model catalog: https://y-api.bestvirtualgoods.com/models
- Integration guide: https://y-api.bestvirtualgoods.com/docs
