Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Qwen · Language model

Qwen3.8 2.4T A95B

qwen/qwen3.8-2.4t-a95b
  • Language
Released Aug 12, 2026

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Specifications

Context window
1,048,576 tokens
Max output
131,072 tokens
Input
text
Output
text
Reasoning effort
xhigh, medium, low (always on)
Tokenizer
Qwen

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • frequency_penalty
  • logit_bias
  • logprobs
  • max_tokens
  • min_p
  • presence_penalty
  • reasoning
  • reasoning_effort
  • repetition_penalty
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_logprobs
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3.8-2.4t-a95b","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'