Captured photo
GLM-5.3
75 +

About

Z.ai's flagship GLM-5.3: hybrid sparse + linear attention, 1.3M-token context, strongest GLM tier for long-horizon coding and agent work. Served via OpenRouter.

Settings

Response length-  The maximum number of tokens to generate in the output.
Temperature-  The temperature of the model. Higher values make the model more creative and lower values make it more focused.
Diversity control-  Top_p. Filters AI responses based on probability.
Lower values = top few likely responses,
Higher values = larger pool of options.
Context length-  The maximum number of tokens to use as input for a model.
Reasoning effort-  How much reasoning the model does before answering. Higher values improve hard tasks but cost more output tokens and time.