Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Z.ai · Language model

GLM 4.6

z-ai/glm-4.6
  • Language
Released Sep 30, 2025

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

Specifications

Context window
204,800 tokens
Max output
16,384 tokens
Input
text
Output
text
Reasoning
supported
Tokenizer
Other
Knowledge cutoff
2025-03-31

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • frequency_penalty
  • logit_bias
  • max_tokens
  • min_p
  • presence_penalty
  • reasoning
  • repetition_penalty
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-4.6","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'