← All models
Google · Language model
Gemini 3.5 Flash
google/gemini-3.5-flash- Language
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Specifications
- Context window
- 1,048,576 tokens
- Max output
- 65,536 tokens
- Input
- text, image, video, file, audio
- Output
- text
- Reasoning effort
- high, medium, low, minimal (always on)
- Tokenizer
- Gemini
- Knowledge cutoff
- 2025-01-01
Supported parameters
Accepted by the model and forwarded as-is by the gateway.
- max_tokens
- reasoning
- reasoning_effort
- response_format
- seed
- stop
- structured_outputs
- temperature
- tool_choice
- tools
- top_p
Use it
POST /api/v1/chat/completions · MCP chat · full reference
curl https://omnirail.org/api/v1/chat/completions \
-H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
-d '{"model":"google/gemini-3.5-flash","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'