Choosing this model
Selection advice by Y-API. The checks below are suggested evaluations, not published test results.
Where to start
Evaluate it for day-to-day pull requests, code explanations and document analysis with explicit acceptance criteria. Keep hard cases for comparison with Opus instead of routing every request to the same tier.
What to watch for
Do not reuse token-count assumptions from an older Claude tokenizer. Nor should you assume every sampling option is accepted: the reviewed parameter list does not include temperature for this entry.
Model specifications
These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.
- Context window
- 1,000,000 tokens
- Input
- Text · Image · File
- Output
- Text
- Listed on OpenRouter
- Parameters listed by OpenRouter
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstool_choicetoolsverbosity
Estimate your cost
- Estimated credit consumed
- $7.00
- Estimated cash equivalent
- $0.70
Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.
$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.
| Model | Estimated credit consumed | Estimated cash equivalent |
|---|---|---|
| Claude Sonnet 5 | $7.00 | $0.70 |
| Claude Opus 5 | $17.50 | $1.75 |
| GLM 5.3 | $3.90 | $0.39 |
Use the API
These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.
A task to try
Review a retry loop that retries every HTTP error three times. Explain which error classes should not be retried and propose bounded backoff behavior.
Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.
curl --fail-with-body --silent --show-error --max-time 120 \
'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
-H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
-H 'Content-Type: application/json' \
--data-binary @- <<'JSON'
{
"model": "anthropic/claude-sonnet-5",
"messages": [
{
"role": "user",
"content": "Review a retry loop that retries every HTTP error three times. Explain which error classes should not be retried and propose bounded backoff behavior."
}
],
"max_tokens": 4096
}
JSONHow to evaluate the result
Distinguish transient failures from authentication and validation errors. For writes, require idempotency before retries; check that the proposal limits both attempts and elapsed time.
Before you choose
Can I copy all of my older Claude parameters?
Do not assume so. Start with the minimal request, then add each needed option. The tokenizer and the accepted parameter set can differ from earlier models or from a native Anthropic endpoint.
Sources & scope
Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.