← All models
OpenAI · Language model
GPT-4.1 Nano
openai/gpt-4.1-nano- Language
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Specifications
- Context window
- 1,047,576 tokens
- Max output
- 32,768 tokens
- Input
- image, text, file
- Output
- text
- Tokenizer
- GPT
- Knowledge cutoff
- 2024-06-30
- Moderation
- provider-moderated
Supported parameters
Accepted by the model and forwarded as-is by the gateway.
- max_completion_tokens
- max_tokens
- response_format
- seed
- structured_outputs
- temperature
- tool_choice
- tools
- top_p
Use it
POST /api/v1/chat/completions · MCP chat · full reference
curl https://omnirail.org/api/v1/chat/completions \
-H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
-d '{"model":"openai/gpt-4.1-nano","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'