cd /news/large-language-models/is-there-a-point-messing-with-strix-… · home topics large-language-models article
[ARTICLE · art-101603] src=forum.level1techs.com ↗ pub= topic=large-language-models verified=true sentiment=· neutral

Is there a point messing with Strix Halo 395 64GB if I have a 5090?

A user asks whether a 64GB Strix Halo 395 machine is useful for local LLM inference alongside an RTX 5090, noting the main advantage is loading models larger than 32GB into memory despite slower inference, and questions if any model better than Qwen 3.8 27B exists for this hardware.

read1 min views1 publishedAug 18, 2026
Is there a point messing with Strix Halo 395 64GB if I have a 5090?
Image: Forum (auto-discovered)

Mainly asking for local LLM use. I have both machines already.

Is there actually anything useful I could do with a 64GB Strix Halo box that I can’t already do better on the 5090?

I’m guessing the main advantage would be models larger than 32GB into memory, even if inference is slower. But does that make it worth playing with one for local LLMs? It seems like there is no model better than Qwen 3.8 27B for this hardware anyway?

── more in #large-language-models 4 stories · sorted by recency
── more on @strix halo 395 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/is-there-a-point-mes…] indexed:0 read:1min 2026-08-18 ·