Model summary
Speed
GLM-5.3-Flash is amongst the leading models in intelligence and well priced when comparing to other models of similar price. The model supports text input, outputs text, and has a 400k tokens context window.
GLM-5.3-Flash scores 57 on the Artificial Analysis Intelligence Index, placing it well above average among comparable models (median: 18). When evaluating the Intelligence Index, it generated 150M tokens, which is very verbose in comparison to the median of 64M.
Pricing for GLM-5.3-Flash is $0.15 per 1M input tokens (competitively priced, median: $0.25) and $0.50 per 1M output tokens (competitively priced, median: $0.90). In total, it cost $138.02 to evaluate GLM-5.3-Flash on the Intelligence Index.
| Reasoning | Yes This page shows the reasoning version of this model. A non-reasoning variant may also exist. |
|---|---|
| Input modality | Supports: text |
| Output modality | Supports: text |
| Context window | 400k ~600 A4 pages of size 12 Arial font |
Metrics are compared against models of the same class:
- Non-reasoning models → compared only with other non-reasoning models
- Reasoning models → compared across both reasoning and non-reasoning
- Open weights models → compared only with other open weights models of the same size class:
- Tiny: ≤4B parameters
- Small: 4B–40B parameters
- Medium: 40B–150B parameters
- Large: >150B parameters
-
Proprietary models → compared across proprietary and open weights models of the same price range, using a blended 3:1 input/output price ratio:
-
<$0.15 per 1M tokens
-
$0.15–$1 per 1M tokens
-
$1 per 1M tokens Highlights
Speed
Intelligence #
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index by Open Weights / Proprietary
Intelligence Evaluations
Agentic real-world work tasks, (Elo-500)/2000
[𝜏³-Banking](/evaluations/tau3-banking)Updated
Agentic tool use
Agentic coding & terminal use
Coding
Humanity's Last ExamUpdated Reasoning & knowledge
Scientific reasoning
Physics reasoning
AA-Omniscience AccuracyUpdated Knowledge
1 - hallucination rate
AA-LCRUpdated Long context reasoning
Agentic knowledge work, Elo
Agentic SaaS workflows
Legal agentic work, criterion pass rate
Agentic business operations
Quantitative analysis on spreadsheets & documents
Instruction following
Long-horizon agentic tasks
Kubernetes incident root-cause analysis
Visual reasoning
AA-Omniscience
AA-Omniscience Index
Intelligence Index Comparisons #
Intelligence Index vs. Cost per Intelligence Index Task
Token Use #
Output Tokens per Intelligence Index Task
Cost #
Cost per Intelligence Index Task
Cost to Run Artificial Analysis Intelligence Index
Pricing: Cache Hit, Input, and Output
Context Window #
Context Window
Frequently Asked Questions #
Common questions about GLM-5.3-Flash GLM-5.3-Flash was released on August 26, 2026.
GLM-5.3-Flash was created by Z AI. GLM-5.3-Flash scores 57 on the Artificial Analysis Intelligence Index, placing it well above average among other reasoning models in a similar price tier (median: 18).
GLM-5.3-Flash costs $0.15 per 1M input tokens (very competitive, median: $0.25) and $0.50 per 1M output tokens (very competitive, median: $0.90), based on Z AI's API.
GLM-5.3-Flash costs $0.15 per 1M input tokens and $0.50 per 1M output tokens (based on Z AI's API). For a blended rate (7:2:1 cache hit/input/output ratio), this is $0.10 per 1M tokens. Pricing may vary by provider. Compare provider pricing
When evaluated on the Intelligence Index, GLM-5.3-Flash generated 150M output tokens, which is at the higher end compared to other reasoning models in a similar price tier (median: 64M).
Yes, GLM-5.3-Flash is a reasoning model. It uses extended thinking or chain-of-thought reasoning to work through complex problems before providing an answer.
GLM-5.3-Flash supports text input.
GLM-5.3-Flash supports text output.
No, GLM-5.3-Flash does not support image input. It can only process text.
No, GLM-5.3-Flash is not multimodal. It only supports text input.
GLM-5.3-Flash has a context window of 400k tokens. This determines how much text and conversation history the model can process in a single request.
No, GLM-5.3-Flash is proprietary. The model weights are not publicly available.
GLM-5.3-Flash is a proprietary model and Z AI has not disclosed the model size or parameter count.
GLM-5.3-Flash achieves a score of 57 on the Artificial Analysis Intelligence Index. This composite benchmark evaluates models across reasoning, knowledge, mathematics, and coding.
Yes, GLM-5.3-Flash is available via API through 1 provider. [Compare API providers](/models/glm-5-3-flash/providers)
GLM-5.3-Flash is available through 1 API provider. [Compare providers](/models/glm-5-3-flash/providers)