{"slug": "new-model-available-glm-5-3-flashx", "title": "New Model Available: GLM 5.3 FlashX", "summary": "Z.ai released GLM-5.3-FlashX, a native multimodal model that delivers inference speeds of up to 200 tokens/s, faster than GLM-5.3-Flash. The model uses a hybrid sparse and linear attention architecture to maintain accurate long-context behavior while reducing compute overhead, and is positioned for efficient coding and long-horizon agent tasks. Pricing on the AI Gateway starts at $0.375 per million tokens, with read costs of $0.075 per million tokens.", "body_md": "GLM-5.3-FlashX is a faster and smoother version of GLM-5.3-Flash, delivering inference speeds of up to 200 tokens/s. GLM-5.3-FlashX is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.\n\n[Back to Models](https://zenmux.ai/models)\n\n## Providers\n\nRoute requests across multiple providers. Copy a provider slug to set your preference.\n\n**$0.375**\n\n*/ M tokens*\n\n**$1.25**\n\n*/ M tokens*\n\n*Read:*\n\n**0.075**/ M tokens\n\n*Write:*\n\n**-**/ M tokens1M--\n\n## Uptime\n\n24hours\nDirect request success rate on AI Gateway and per-provider.\n\n## Throughput\n\n24hours\nP50 throughput on live AI Gateway traffic, in tokens per second (TPS).\n\n## Latency\n\n24hours\nP50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.\n\n## Activity\n\nToken volume and request traffic to this model over time.\n\n## Benchmarks\n\nScores on standardized evaluations. Higher percentages are better — and rank percentile shows\n\nMetrics sourced from[Artificial Analysis](https://artificialanalysis.ai/)\n\n## Apps\n\nPublic apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. [View All](https://zenmux.ai/analytics/apps)\n\n## Related Models\n\nMore models from [Z.ai](https://zenmux.ai/z-ai)", "url": "https://wpnews.pro/news/new-model-available-glm-5-3-flashx", "canonical_source": "https://zenmux.ai/z-ai/glm-5.3-flashx", "published_at": "2026-09-18 06:01:47+00:00", "updated_at": "2026-09-18 06:25:17.934609+00:00", "lang": "en", "topics": ["large-language-models", "ai-products", "ai-agents", "ai-infrastructure"], "entities": ["Z.ai", "GLM-5.3-FlashX", "GLM-5.3-Flash", "Artificial Analysis", "AI Gateway"], "alternates": {"html": "https://wpnews.pro/news/new-model-available-glm-5-3-flashx", "markdown": "https://wpnews.pro/news/new-model-available-glm-5-3-flashx.md", "text": "https://wpnews.pro/news/new-model-available-glm-5-3-flashx.txt", "jsonld": "https://wpnews.pro/news/new-model-available-glm-5-3-flashx.jsonld"}}