Choosing this model
Selection advice by Y-API. The checks below are suggested evaluations, not published test results.
Where to start
Evaluate it on agent workloads that run unattended for extended periods, where the model has to verify its own intermediate results rather than waiting for a human to check each step.
What to watch for
Reasoning effort is exposed but its effect on cost is not something this page can predict — reasoning tokens are billed, so a higher setting can change the bill substantially. Closed weights mean there is no card to inspect.
Model specifications
These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.
- Context window
- 1,000,000 tokens
- Input
- Text · Image · File
- Output
- Text
- Listed on OpenRouter
- Parameters listed by OpenRouter
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstool_choicetoolsverbosity
API pricing & cost estimate
- Estimated credit consumed
- $17.50
- Estimated cash equivalent
- $1.75
Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.
$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.
| Model | Estimated credit consumed | Estimated cash equivalent |
|---|---|---|
| Claude Opus 4.7 | $17.50 | $1.75 |
| Claude Opus 4.6 | $17.50 | $1.75 |
| Claude Sonnet 5 | $7.00 | $0.70 |
API integration examples
These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.
A task to try
You are running a long migration unattended. State what you would check after each batch before continuing, and what condition should make you stop and report instead.
Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.
curl --fail-with-body --silent --show-error --max-time 120 \
'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
-H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
-H 'Content-Type: application/json' \
--data-binary @- <<'JSON'
{
"model": "anthropic/claude-opus-4.7",
"messages": [
{
"role": "user",
"content": "You are running a long migration unattended. State what you would check after each batch before continuing, and what condition should make you stop and report instead."
}
],
"max_tokens": 4096
}
JSONHow to evaluate the result
A good answer defines a concrete, checkable condition for both continuing and stopping. Vague instructions like "verify correctness" do not count as a stop condition.
Before you choose
How does 4.7 differ from 4.6 in practice?
Both list the same window, modalities and price. The documented difference is the generation and its agentic tuning, so compare them on your own long-running tasks rather than assuming the newer one wins.
Sources & scope
Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.