Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Google · Language model

Gemini 3.8 Flash

google/gemini-3.8-flash
  • Language
Released Sep 2, 2026

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Specifications

Context window
1,048,576 tokens
Max output
65,536 tokens
Input
text, image, video, file, audio
Output
text
Reasoning effort
high, medium, low (always on)
Tokenizer
Gemini

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • max_tokens
  • reasoning
  • reasoning_effort
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.8-flash","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'