Captured photo
Gemini 3.8 Flash
60 +

Percs

Long context

About

Google's most intelligent Flash-tier model, GA 2026-09-02: significant gains over 3.7 Flash on software engineering, agentic tasks and multi-step reasoning. 1.05M-token context. Served via OpenRouter.

Available controls: Response length, Temperature, Diversity control, Context length, Reasoning effort.

Settings

Response length-  The maximum number of tokens to generate in the output.
Temperature-  The temperature of the model. Higher values make the model more creative and lower values make it more focused.
Diversity control-  Top_p. Filters AI responses based on probability.
Context length-  The maximum number of tokens to use as input for a model.
Reasoning effort-  How much reasoning the model does before answering. Reasoning is mandatory for this model. Higher values improve hard tasks but cost more output tokens and time.