# The fastest and cheapest GLM 5.3 Flash endpoint

> Source: <https://runinfra.ai/inference-api/glm-5-3-flash>
> Published: 2026-08-26 20:30:40+00:00

`zai-org/GLM-5.3-Flash`

GLM 5.3 Flash is an LLM listed in RunInfra Model APIs. RunInfra serves it as zai-org/GLM-5.3-Flash at $0.10 per 1M input tokens and $0.40 per 1M output tokens. Its context window is 1,048,576 tokens. The API provides OpenAI-compatible chat completions.

USD, pay per token

Confirm how your client reaches this model.

Check the limits your workload must fit.

See which request modes the API supports.

Set RUNINFRA_GATEWAY_KEY to your workspace API key before using an example.

Verify the company and operating credentials behind this API.

© 2026 RunInfra. All rights reserved.
