GLM 5.3 Flash is a Z.ai model tracked in Sim. It supports a 1M token context window. Pricing starts at $0.15/1M input tokens and $0.5/1M output tokens. Key capabilities include Temperature 0-1, Tool choice, Structured outputs. Best for long-context retrieval, large documents, and high-memory workflows.
Best forBest for long-context retrieval, large documents, and high-memory workflows.
Temperature0 to 1
Reasoning effortNot supported
VerbosityNot supported
Thinking levelsNot supported
Structured outputsSupported
Tool choiceSupported
Computer useNot supported
Deep researchNot supported
Memory supportSupported
Max output tokens131k
Frequently asked questions
GLM 5.3 Flash is a Z.ai model available in Sim. GLM 5.3 Flash is a Z.ai model tracked in Sim. It supports a 1M token context window. Pricing starts at $0.15/1M input tokens and $0.5/1M output tokens. Key capabilities include Temperature 0-1, Tool choice, Structured outputs.
GLM 5.3 Flash is listed at $0.15/1M input tokens, and $0.5/1M output tokens.
GLM 5.3 Flash supports a context window of 1M tokens in Sim. In an agent, this determines how much conversation history, tool outputs, and retrieved documents the model can hold in a single call.
GLM 5.3 Flash supports the following capabilities in Sim: Temperature 0-1, Tool choice, Structured outputs, Max output 131k.
Best for long-context retrieval, large documents, and high-memory workflows. When used in a Sim workflow, it can be selected in any Agent block from the model picker.