{"slug": "new-model-available-z-ai-glm-5-3-flash", "title": "New Model Available: Z.AI: GLM 5.3 Flash", "summary": "Z.ai released GLM-5.3-Flash, a native multimodal model designed for efficient coding and long-horizon agent tasks, featuring a hybrid sparse and linear attention architecture that reduces compute overhead while maintaining long-context accuracy. Pricing starts at $0.15 per million input tokens and $0.5 per million output tokens, with some providers offering discounted rates.", "body_md": "GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.\n\n[Back to Models](https://zenmux.ai/models)\n\n## Providers\n\nRoute requests across multiple providers. Copy a provider slug to set your preference.\n\n**$0.15**\n\n*/ M tokens*\n\n**$0.5**\n\n*/ M tokens*\n\n*Read:*\n\n**0.03**/ M tokens\n\n*Write:*\n\n**-**/ M tokens1M6.3s37.6tps\n\n**$0.15**\n\n*/ M tokens*\n\n**$0.5**\n\n*/ M tokens*\n\n*Read:*\n\n**0.03**/ M tokens\n\n*Write:*\n\n**-**/ M tokens1M1.96s69.9tps\n\n~~$0.15~~\n\n**$0.075**\n\n*/ M tokens*\n\n~~$0.5~~\n\n**$0.25**\n\n*/ M tokens*\n\n*Read:*\n\n~~0.03~~\n\n**0.015**/ M tokens\n\n*Write:*\n\n**-**/ M tokens1M5.64s23.6tps\n\n## Uptime\n\n24hours\nDirect request success rate on AI Gateway and per-provider.\n\n## Throughput\n\n24hours\nP50 throughput on live AI Gateway traffic, in tokens per second (TPS).\n\n## Latency\n\n24hours\nP50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.\n\n## Activity\n\nToken volume and request traffic to this model over time.\n\n## Apps\n\nPublic apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. [View All](https://zenmux.ai/analytics/apps)\n\n## Related Models\n\nMore models from [Z.ai](https://zenmux.ai/z-ai)", "url": "https://wpnews.pro/news/new-model-available-z-ai-glm-5-3-flash", "canonical_source": "https://zenmux.ai/z-ai/glm-5.3-flash", "published_at": "2026-08-26 15:12:37+00:00", "updated_at": "2026-09-09 06:57:57.826532+00:00", "lang": "en", "topics": ["large-language-models", "generative-ai", "ai-products"], "entities": ["Z.ai", "GLM-5.3-Flash"], "also_reported_by": [], "alternates": {"html": "https://wpnews.pro/news/new-model-available-z-ai-glm-5-3-flash", "markdown": "https://wpnews.pro/news/new-model-available-z-ai-glm-5-3-flash.md", "text": "https://wpnews.pro/news/new-model-available-z-ai-glm-5-3-flash.txt", "jsonld": "https://wpnews.pro/news/new-model-available-z-ai-glm-5-3-flash.jsonld"}}