Choosing this model
Selection advice by Y-API. The checks below are suggested evaluations, not published test results.
Where to start
Try it for document extraction with explicit missing-information handling. Make it distinguish a supplied fact from an assumption, especially when a short business request omits a date or amount.
What to watch for
A published emphasis on grounded answers is not evidence that hallucination has been eliminated. Include underspecified questions in your evaluation and check how reasoning settings reach the gateway.
Model specifications
These are OpenRouter model-level specifications, not a Y-API compatibility test. A particular route may accept fewer input formats, parameters or tokens. Verify the features you need with a small request; a successful text response does not validate vision or tool calling.
- Context window
- 262,144 tokens
- Input
- Text
- Output
- Text
- Listed on OpenRouter
- Parameters listed by OpenRouter
frequency_penaltyinclude_reasoninglogit_biasmax_completion_tokensmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Estimate your cost
The current catalog marks this model as free: calls do not deduct credit. This does not promise unlimited capacity, permanent availability or a future free price.
- Estimated credit consumed
- Free — no credit deducted
- Estimated cash equivalent
- Free — no credit deducted
Illustrative token budget, not measured usage or a quote. Includes input and output; excludes separate media charges, cache discounts and retries. Include billed reasoning tokens in your output budget where applicable. Actual billing follows usage returned by the service.
$1 paid adds $10 of credit. Cash equivalent = credit consumed ÷ 10; this is not the model publisher’s list price.
| Model | Estimated credit consumed | Estimated cash equivalent |
|---|---|---|
| Hy3 | Free — no credit deducted | |
| Hy4 preview | $2.50 | $0.25 |
| MiniMax M2.7 | Free — no credit deducted | |
Use the API
These minimal, text-only requests use the exact Y-API model ID. They do not demonstrate image, audio, video, file or tool support. The output limit also needs room for reasoning; an empty answer with finish_reason=length can mean the budget was exhausted.
A task to try
From this note extract supplier, amount and due_date; use null when absent: Pay Acme Tools USD 240 for the replacement part.
Set the TOKEN environment variable to a key from your console. Run this on your server or locally; never expose a key in browser code or a public repository.
curl --fail-with-body --silent --show-error --max-time 120 \
'https://api.y-api.bestvirtualgoods.com/v1/chat/completions' \
-H "Authorization: Bearer ${TOKEN:?Set TOKEN first}" \
-H 'Content-Type: application/json' \
--data-binary @- <<'JSON'
{
"model": "tencent/hy3",
"messages": [
{
"role": "user",
"content": "From this note extract supplier, amount and due_date; use null when absent: Pay Acme Tools USD 240 for the replacement part."
}
],
"max_tokens": 4096
}
JSONHow to evaluate the result
Expect the supplied supplier and amount, with a null due date. Add contradictory notes and follow-up corrections to see whether the model preserves the latest constraint.
Before you choose
Does Hy3 always use a reasoning mode?
OpenRouter describes a direct no-thinking default plus selectable reasoning modes. Do not assume the Y-API route preserves that default; verify the response and token usage with your intended settings.
Sources & scope
Technical facts come from the cited model page and, where available, its linked publisher card. A card describes that checkpoint; it is not proof of the weights a gateway serves. Prices come only from the Y-API catalog. We do not claim measured latency, uptime or benchmark scores for this endpoint.