Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Qwen · Language model

Qwen3.5-Flash

qwen/qwen3.5-flash-02-23
  • Language
Released Feb 25, 2026

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Specifications

Context window
1,000,000 tokens
Max output
65,536 tokens
Input
text, image, video
Output
text
Reasoning
supported
Tokenizer
Qwen3

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • frequency_penalty
  • max_tokens
  • presence_penalty
  • reasoning
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3.5-flash-02-23","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'