cd/entity/Qwen· home› entities› Qwen
grep -l @qwen /news/*.json | wc -l → 940

Qwen

mentions 940 type Organization page 10/47 feed RSS

// recent coverage 940 mentions

15:05
2026-09-01
dev.to
artificial-intelligence

Running Local AI Models on a Consumer GPU: A 2026 Field Test

A developer's field test of running local AI models on consumer GPUs in 2026 finds that memory capacity is the key constraint, with 8GB handling 3B-4B models and 32GB enabling 14B-30B models via quant…

11:34
2026-08-31
dev.to
developer-tools

Why I Stopped Using JSON Tool-Calling for My Coding Agent

A developer built CodePilot, an open-source, embeddable Python runtime for coding agents, after finding that JSON tool-calling broke on large code payloads. The runtime uses a plain-text search-and-re…

21:45
2026-08-30
thewatershed.markpesce.com
artificial-intelligence

Century

In August, Chinese AI labs released a series of frontier and near-frontier open-weights models, including DeepSeek-v4-Flash-0731, Alibaba's Qwen3.8-27B, Z.ai's GLM 5.3 and GLM 5.3 Flash, and Qwen3.8-F…

00:07
2026-08-30
dev.to
ai-infrastructure

Standing Up a GPU Cluster on AKS for vLLM

Josef Doornink, an engineer, published a guide to standing up a GPU cluster on Azure Kubernetes Service (AKS) for serving vLLM models. The walkthrough covers requesting GPU quota, creating a cluster w…

08:02
2026-08-29
dev.to
large-language-models

Why I Stopped Chasing the Newest LLM (And What I Run Instead)

A developer has stopped chasing the newest large language models and instead runs a fixed stack of six models across three machines, arguing that frequent model churn costs more time than it saves. Th…

17:31
2026-08-28
blog.brokk.ai
artificial-intelligence

A Coding Subscription Tier List

A new informal tier list from Brokk, based on coding performance in Rust static analysis tooling, ranks OpenAI's Luna subscription as the top value, solving more than 10x as many DeepSWE tasks per $20…

14:28
2026-08-28
notes.npilk.com
artificial-intelligence

Free intelligence is getting better

Open-source and free AI models are becoming nearly as capable as frontier American models, with local models like Ornith 1.5 9B running on consumer hardware and free API models from Nvidia and others …

00:00
2026-08-28
mindstudio.ai
ai-infrastructure

Dark Bloom: Rent Out Your Mac for AI Inference and Get Paid

Dark Bloom, a distributed inference network, pays Mac owners to share idle compute for AI inference, with payouts via Stripe and models served through OpenRouter at lower prices. The project, which re…

16:25
2026-08-27
forum.level1techs.com
large-language-models

Llama.cpp Qwen devolves into repeating /'s

A user reports that Gemma 4:26B, a 26-billion-parameter mixture-of-experts model, begins repeating tokens and entering loops as context length approaches 80,000 tokens, with severe degradation above t…

← prev page 10 / 47 next →
// co-occurs with top 8 entities
// topics top 6 topics