cd /news/large-language-models/glm-5-3-flash-unusably-slow · home topics large-language-models article
[ARTICLE · art-132976] src=kagifeedback.org ↗ pub= topic=large-language-models verified=true sentiment=↓ negative

GLM 5.3 Flash unusably slow

Users of GLM 5.3 Flash reported the model has become unusably slow, with response times of at least 3 minutes for simple questions, according to complaints posted on a discussion thread. Commenter shurik said the delay occurs when using the Kagi assistant app on iOS with GLM 5.3 Flash, while comicfrieze reported similar slowness and timeouts with QWEN 3.8 27b and DeepSeek V4.1 Flash over roughly the past two days, dating the issue to around September 15-16.

by read1 min views2 publishedSep 17, 2026

spiffytech Recently GLM 5.3 Flash has gone from usable to now it's so slow I just give up and walk away for minutes waiting for it to answer simple questions. Model responds quickly enough that I feel comfortable waiting on it.

shurik same here - it is painfully slow. takes at least 3 (!) minutes to answer a simple question when using the kagi assistant app on ios with GLM 5.3 Flash.

comicfrieze It's not just GLM 5.3, but I've confirmed using QWEN 3.8 27b and DeepSeek V4.1 Flash. It's painfully slow, if it doesn't time out completely. I've noticed this over the past 2 days or so (definitely September 16... maybe September 15).

── more in #large-language-models 4 stories · sorted by recency
── more on @glm 5.3 flash 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/glm-5-3-flash-unusab…] indexed:0 read:1min 2026-09-17 ·