Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Google · Language model

Gemini 3.5 Flash Lite

google/gemini-3.5-flash-lite
  • Language
Released Jul 21, 2026

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Specifications

Context window
1,048,576 tokens
Max output
65,536 tokens
Input
text, image, video, file, audio
Output
text
Reasoning effort
high, medium, low, minimal (always on)
Tokenizer
Gemini

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • max_tokens
  • reasoning
  • reasoning_effort
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.5-flash-lite","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'