cd/entity/Qwen· home› entities› Qwen
grep -l @qwen /news/*.json | wc -l → 939

Qwen

mentions 939 type Organization page 8/47 feed RSS

// recent coverage 939 mentions

16:27
2026-09-10
narilabs.com
ai-products

10x Cheaper TTS at 50ms Time-to-First-Audio

Nari Labs launched a Free Public Beta of realtime text-to-speech endpoints built on its own inference engine for the Qwen3-TTS model, claiming 50 ms time to first audio in Fast mode versus 300 ms for …

18:48
2026-09-09
promptcube3.com
developer-tools

Solving the Qwen Coder local setup lag in VS Code

A developer reports that running Qwen2.5-Coder-32B via Ollama in VS Code caused 2-3 second latency, fixed by switching to the q4_K_M quantized version and setting num_ctx to 16384, improving response …

21:01
2026-09-08
dev.to
artificial-intelligence

Open Models Are Everywhere. But What Is an Open Stack?

Developer Rijul, creator of LiveReview, explains the concept of an 'open stack' in AI, contrasting it with open-weight models. He notes that open-weight models like Llama, Mistral, Gemma, and Qwen pro…

13:04
2026-09-08
news.ycombinator.com
developer-tools

Why human syntax breaks LLMs (and how to fix agentic coding)

AgentScript (ASL), an open-source, statically typed language using single-pass S-expressions that compiles to native Rust, Go, TypeScript, and WebAssembly, reduces syntax repair waste in coding agents…

12:08
2026-09-08
sourcefeed.dev
ai-tools

Why 28,000 developers told their agent they have ADHD

A GitHub skill called i-have-adhd, created by user ayghri, has amassed 28,000 stars and 1,800 forks since May by instructing coding agents to format responses for readers with ADHD, leading with next …

06:52
2026-09-08
mmoustafa.com
ai-infrastructure

So you want to use OpenRouter?

OpenRouter's multi-provider routing causes significant performance variability for the same AI model, according to a developer who runs the iMessage assistant Olly and has processed over 18 million me…

02:35
2026-09-08
dev.to
large-language-models

Qwen4 Isn’t Here Yet, but Qwen3.8-Flash-Next Tells Us a Lot

Qwen3.8-Flash-Next, a sparse Mixture-of-Experts model with roughly 125B total parameters but only about 6B active per token, offers a preview of Qwen's architectural direction for Qwen4. The model emp…

15:07
2026-09-07
dev.to
artificial-intelligence

The self-hosting math for LLMs quietly inverted this year

A developer's analysis of self-hosting economics for AI models reveals that the cost calculus has shifted dramatically in 2025, with API token prices falling while GPU hardware costs have risen sharpl…

← prev page 8 / 47 next →
// co-occurs with top 8 entities
// topics top 6 topics