cd /news/large-language-models/trying-to-improve-output-speed-with-… · home topics large-language-models article
[ARTICLE · art-138295] src=discuss.huggingface.co ↗ pub= topic=large-language-models verified=true sentiment=· neutral

Trying to improve output speed with MLX models on LM Studio

An LM Studio user running version 0.4.25+1 on a MacBook Pro M5 Max with 128 GB of RAM reports getting 15-17 tokens per second from the mlx-community/Qwen3.8-27B-8bit model and cannot enable speculative decoding because LM Studio returns "No compatible draft models found for your current model selection" after downloading mlx-community/Qwen3.8-27B-MTP-8bit as a draft model. The user is running the LM Studio MLX (Apple M5) v1.11.0 engine and is asking for an alternative draft model or other ways to improve output speed.

read1 min views3 publishedSep 23, 2026

Hello, I’ve been using LM Studio on my Macbook Pro M5 Max (128 GB) with mlx-community/qwen3.8-27b without issues, so I’m looking for ways to improve output speed (right now I’m getting 15-17 t/s) and found about speculative decoding. Looking around I found mlx-community/Qwen3.8-27B-MTP-8bit is a draft model for that, so I downloaded it, but when I try to enable speculative decoding, LM Studio says “No compatible draft models found for your current model selection”.

My LM Studio version is 0.4.25+1 and the MLX engine I’m using is LM Studio MLX (Apple M5) v1.11.0.

If anyone can suggest another draft model or any other way to improve output speed I’d really appreciate it. Thanks in advance.

── more in #large-language-models 4 stories · sorted by recency
── more on @lm studio 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/trying-to-improve-ou…] indexed:0 read:1min 2026-09-23 ·