Qwen3.8 27B addition in words
Simon Willison reran Colin Frasier's GPT-4o "sum in words" experiment on local hardware using Qwen3.8-27B-Q4_K_M.gguf on an Nvidia DGX Spark, with reasoning enabled the model answered 167 of 169 one-s…
Simon Willison reran Colin Frasier's GPT-4o "sum in words" experiment on local hardware using Qwen3.8-27B-Q4_K_M.gguf on an Nvidia DGX Spark, with reasoning enabled the model answered 167 of 169 one-s…
TensorFold 0.3.6.2 delivered decode speeds over 62 tokens per second on a single stream and 119 tokens per second across five concurrent streams running Qwen3.8-Flash-Next on a single Nvidia DGX Spark…
NextLM Inc. launched its AI Prospecting Agent on Google LLC's Cloud Marketplace and Gemini Enterprise, letting customers buy the service with existing Google Cloud spending commitments and invoke the …
MSI's EdgeXpert, an Nvidia DGX Spark with 128GB of RAM, can remotely control an MSI Cubi NUC AI+ running Windows 11 to launch programs and run tools, according to a Level1Techs demonstration. The setu…
Apple's M5 Ultra Mac Studio and M6 Mac mini extend the company's lead in consumer and professional AI hardware, according to a first-look test by Jason Hiner, who reports the M5 Ultra Mac Studio carri…
A user running hybrid GPU+CPU inference on 4x Nvidia V100 GPUs reported beating an Nvidia RTX 5090 in token generation (TG) for Qwen 3.8 27B and outperforming two DGX Spark systems with Qwen 3.8 Flash…
OpenFaaS Ltd purchased 4 Nvidia DGX Spark systems, connecting the first two with a high-speed Connect-X cable, to run local AI models including Qwen 3.5-3.8 27B and larger models such as DeepSeek and …
Jeff Morgan, co-founder and CEO of Ollama, said open AI models are gaining momentum as enterprises and developers seek more control, cost savings, and privacy, with Ollama now used across 80% of the F…
Perplexity launched Portable Computer, an on-device AI offering that runs the Nvidia DGX Spark with Qwen 3.8 27B or PPLX 27B, keeping data local and charging only when tasks escalate to the cloud. The…
Perplexity has unveiled Portable Computer, a new version of its Personal Computer agent that runs AI models locally on Linux systems, offering faster performance, enhanced security, and no inference f…
AGI Bar in Beijing's Zhongguancun technology district offers free DeepSeek-powered coding tokens to customers, blending AI with nightlife, powered by two Nvidia DGX Spark machines. Owner Song De, an i…
AGI Bar, a Beijing bar opened by independent AI developer Song De in the Zhongguancun tech hub, offers free DeepSeek tokens with every pint, powered by two Nvidia DGX Spark computers, and serves a sig…
DeepSeek V4 Flash, a post-trained update to DeepSeek's existing V4 Flash preview model with 284 billion parameters, jumped from 7% to 54% on the DeepSweep agentic coding benchmark, rivaling larger mod…
Mia's AI Lab published an open-source deployment stack on July 23rd that runs Z.ai's 753-billion-parameter GLM-5.2 model across three Nvidia DGX Spark computers, creating a desktop cluster with a 248,…
Poolside released Laguna S 2.1 on July 21st, an open-weight 118-billion-parameter Mixture-of-Experts coding model designed to run long-running software agents on user-controlled hardware. The model ac…
AMD has listed a Strix Halo mini PC powered by the Ryzen AI Max+ 395 APU for $3,999, available starting July 10 through MicroCenter. The dev-kit system features 128 GB of shared memory and is designed…