Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Qwen · Language model

Qwen3 Max Thinking

qwen/qwen3-max-thinking
  • Language
Released Feb 9, 2026

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it...

Specifications

Context window
262,144 tokens
Max output
65,536 tokens
Input
text
Output
text
Reasoning
supported
Tokenizer
Qwen

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • frequency_penalty
  • logprobs
  • max_tokens
  • presence_penalty
  • reasoning
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_logprobs
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-max-thinking","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'