Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

OpenAI · Language model

GPT-4.1 Nano

openai/gpt-4.1-nano
  • Language
Released Apr 14, 2025

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

Specifications

Context window
1,047,576 tokens
Max output
32,768 tokens
Input
image, text, file
Output
text
Tokenizer
GPT
Knowledge cutoff
2024-06-30
Moderation
provider-moderated

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • max_completion_tokens
  • max_tokens
  • response_format
  • seed
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"openai/gpt-4.1-nano","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'