cd/entity/AMD Strix Halo· home› entities› AMD Strix Halo
grep -l @amd strix halo /news/*.json | wc -l → 14

AMD Strix Halo

mentions 14 type Person feed RSS

// recent coverage 14 mentions

10:59
2026-10-07
tangled.org
large-language-models

Fast single-box DeepSeek v4.1 Flash runtime

A new recipe for running DeepSeek V4.1 Flash on a single AMD Strix Halo machine (gfx1151 GPU, 128 GB unified memory, several NVMe drives) reports roughly 450 tokens per second prefill and 15 tokens pe…

18:14
2026-09-30
forum.level1techs.com
ai-infrastructure

Ryzen AI Halo: Halogen Server Testing Notes

A community tester published deployment notes for the Halogen Flash Server, an optimized server for running Qwen3.8-Flash-Next on AMD Strix Halo (gfx1151) hardware, reporting a roughly 10% performance…

17:40
2026-08-30
forum.level1techs.com
artificial-intelligence

So, like, how do I measure my local LLM environment and expand it?

A developer running a local LLM environment with an RTX 5090 32GB, AMD Strix Halo 128GB, and RTX 6000 Pro Blackwell 96GB reports that CPU and RAM are nearly irrelevant for fully offloaded workloads, w…

11:48
2026-08-30
forum.level1techs.com
large-language-models

Kimi k3 locally with 8gb ram 0 vram?!?!

A new C99 program under 1 MB enables running Kimi k3, a large language model requiring nearly 2 TB of storage, on systems with as little as 8 GB RAM and no GPU, though at slow speeds. Users report ach…

14:25
2026-08-28
forum.level1techs.com
artificial-intelligence

Strix Halo & external GPU

A Linux novice reports that combining AMD's Strix Halo integrated GPU (Radeon 8060S) with an external Nvidia RTX 2080 Ti via Oculink for local LLM inference is problematic, with LM Studio failing to d…

22:54
2026-08-20
forum.level1techs.com
ai-infrastructure

GMKtec EVO-X3 AI Workstation: Lemonade, Benchmarks, and specs

GMKtec's EVO-X3 AI workstation, featuring AMD Strix Halo, is shipping with delays and questionable VAT documentation, according to forum user ewook on Level1Techs. The user also questions whether addi…

23:39
2026-06-30
latent.space
large-language-models

Ahmad Osman on why local AI is catching up

Ahmad Osman, founder of Osmantic, argued at the AI Engineer World's Fair that local AI is rapidly catching up to proprietary frontier models, driven by shrinking gaps in open-source LLMs and improved …

13:44
2026-06-30
aimultiple.com
ai-products

DGX Spark vs. Mac Studio and Halo

NVIDIA's DGX Spark, a $4,699 desktop AI supercomputer with 128GB unified memory, launched in 2025, offering one petaflop of FP4 performance. Benchmarks show it excels at prompt processing but lags in …

// co-occurs with top 8 entities
// topics top 6 topics