Gemini 3.8 Flash
60 +
Percs
Long context
About
Google's most intelligent Flash-tier model, GA 2026-09-02: significant gains over 3.7 Flash on software engineering, agentic tasks and multi-step reasoning. 1.05M-token context. Served via OpenRouter.
Available controls: Response length, Temperature, Diversity control, Context length, Reasoning effort.
Settings
Response length- The maximum number of tokens to generate in the output.
Temperature- The temperature of the model. Higher values make the model more creative and lower values make it more focused.
Diversity control- Top_p. Filters AI responses based on probability.
Context length- The maximum number of tokens to use as input for a model.
Reasoning effort- How much reasoning the model does before answering. Reasoning is mandatory for this model. Higher values improve hard tasks but cost more output tokens and time.