# GLM-5.3-Flash Intelligence, Performance and Price Analysis

> Source: <https://artificialanalysis.ai/models/glm-5-3-flash>
> Published: 2026-08-26 14:58:54+00:00

# GLM-5.3-Flash Intelligence, Performance & Price Analysis

### Model summary

#### Speed

GLM-5.3-Flash is amongst the leading models in intelligence and well priced when comparing to other models of similar price. The model supports text input, outputs text, and has a 400k tokens context window.

GLM-5.3-Flash scores 57 on the Artificial Analysis Intelligence Index, placing it well above average among comparable models (median: 18). When evaluating the Intelligence Index, it generated 150M tokens, which is very verbose in comparison to the median of 64M.

Pricing for GLM-5.3-Flash is $0.15 per 1M input tokens (competitively priced, median: $0.25) and $0.50 per 1M output tokens (competitively priced, median: $0.90). In total, it cost $138.02 to evaluate GLM-5.3-Flash on the Intelligence Index.

| Reasoning | Yes This page shows the reasoning version of this model. A non-reasoning variant may also exist. |
|---|---|
| Input modality | Supports: text |
| Output modality | Supports: text |
| Context window | 400k ~600 A4 pages of size 12 Arial font |

Metrics are compared against models of the same class:

- Non-reasoning models → compared only with other non-reasoning models
- Reasoning models → compared across both reasoning and non-reasoning
- Open weights models → compared only with other open weights models of the same size class:
- Tiny: ≤4B parameters
- Small: 4B–40B parameters
- Medium: 40B–150B parameters
- Large: >150B parameters
- Proprietary models → compared across proprietary and open weights models of the same price range, using a blended 3:1 input/output price ratio:
- <$0.15 per 1M tokens
- $0.15–$1 per 1M tokens
- >$1 per 1M tokens

Highlights

### Speed

## Intelligence

### Artificial Analysis Intelligence Index

### Artificial Analysis Intelligence Index by Open Weights / Proprietary

### Intelligence Evaluations

Agentic real-world work tasks, (Elo-500)/2000

[𝜏³-Banking](/evaluations/tau3-banking)Updated

Agentic tool use

Agentic coding & terminal use

Coding

[Humanity's Last Exam](/evaluations/humanitys-last-exam)Updated

Reasoning & knowledge

Scientific reasoning

Physics reasoning

[AA-Omniscience Accuracy](/evaluations/omniscience)Updated

Knowledge

1 - hallucination rate

[AA-LCR](/evaluations/artificial-analysis-long-context-reasoning)Updated

Long context reasoning

Agentic knowledge work, Elo

Agentic SaaS workflows

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

Instruction following

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

### AA-Omniscience

### AA-Omniscience Index

## Intelligence Index Comparisons

### Intelligence Index vs. Cost per Intelligence Index Task

## Token Use

### Output Tokens per Intelligence Index Task

## Cost

### Cost per Intelligence Index Task

### Cost to Run Artificial Analysis Intelligence Index

### Pricing: Cache Hit, Input, and Output

## Context Window

### Context Window

## Frequently Asked Questions

Common questions about GLM-5.3-Flash

GLM-5.3-Flash was released on August 26, 2026.

GLM-5.3-Flash was created by Z AI.

GLM-5.3-Flash scores 57 on the Artificial Analysis Intelligence Index, placing it well above average among other reasoning models in a similar price tier (median: 18).

GLM-5.3-Flash costs $0.15 per 1M input tokens (very competitive, median: $0.25) and $0.50 per 1M output tokens (very competitive, median: $0.90), based on Z AI's API.

GLM-5.3-Flash costs $0.15 per 1M input tokens and $0.50 per 1M output tokens (based on Z AI's API). For a blended rate (7:2:1 cache hit/input/output ratio), this is $0.10 per 1M tokens. Pricing may vary by provider. [Compare provider pricing](/models/glm-5-3-flash/providers)

When evaluated on the Intelligence Index, GLM-5.3-Flash generated 150M output tokens, which is at the higher end compared to other reasoning models in a similar price tier (median: 64M).

Yes, GLM-5.3-Flash is a reasoning model. It uses extended thinking or chain-of-thought reasoning to work through complex problems before providing an answer.

GLM-5.3-Flash supports text input.

GLM-5.3-Flash supports text output.

No, GLM-5.3-Flash does not support image input. It can only process text.

No, GLM-5.3-Flash is not multimodal. It only supports text input.

GLM-5.3-Flash has a context window of 400k tokens. This determines how much text and conversation history the model can process in a single request.

No, GLM-5.3-Flash is proprietary. The model weights are not publicly available.

GLM-5.3-Flash is a proprietary model and Z AI has not disclosed the model size or parameter count.

GLM-5.3-Flash achieves a score of 57 on the Artificial Analysis Intelligence Index. This composite benchmark evaluates models across reasoning, knowledge, mathematics, and coding.

Yes, GLM-5.3-Flash is available via API through 1 provider. [Compare API providers](/models/glm-5-3-flash/providers)

GLM-5.3-Flash is available through 1 API provider. [Compare providers](/models/glm-5-3-flash/providers)
