GLM-5.3-Flash
5 +
About
Zhipu AI's 320B-A18B natively multimodal MoE with a 1M token context window, supersedes the GLM-5.2 base tier.
Settings
Response length- The maximum number of tokens to generate in the output.
Temperature- The temperature of the model. Higher values make the model more creative and lower values make it more focused.
Diversity control- Top_p. Filters AI responses based on probability.
Context length- The maximum number of tokens to use as input for a model.
