cd/entity/Qwen2.5-Coder-32B· home entities Qwen2.5-Coder-32B
grep -l @qwen2.5-coder-32b /news/*.json | wc -l → 6

Qwen2.5-Coder-32B

mentions 6 type Organization feed RSS

// recent coverage 6 mentions

22:05
2026-09-12
promptcube3.com
large-language-models

Gemini Forum, Qwen Coder local setup

A developer resolved CUDA out-of-memory crashes running Qwen2.5-Coder-32B on a 24GB RTX 3090 by capping the context window at 8k tokens instead of 32k via a custom Ollama Modelfile, cutting latency fr…

18:48
2026-09-09
promptcube3.com
developer-tools

Solving the Qwen Coder local setup lag in VS Code

A developer reports that running Qwen2.5-Coder-32B via Ollama in VS Code caused 2-3 second latency, fixed by switching to the q4_K_M quantized version and setting num_ctx to 16384, improving response …

13:31
2026-09-08
promptcube3.com
large-language-models

Qwen Coder local setup

A developer's guide to setting up Qwen2.5-Coder locally recommends using the Q4_K_M GGUF quantization to run the 32B model on 16GB machines, reducing RAM requirements from 64GB+ to 20GB with minimal i…

07:18
2026-08-21
vgel.me
artificial-intelligence

Small Models Can Introspect, Too (2025)

A researcher at Alignment of Complex Systems showed that a 32B open-source model, Qwen2.5-Coder-32B, can subtly introspect when external concepts are injected into its activations, despite appearing u…

00:00
2026-08-02
zackproser.com
large-language-models

Sampling Inkling and the Alias Pattern

Thinking Machines' Inkling, a 975B total / 41B active MoE model with Apache 2.0 license and 1M context, is being trialed via a bash wrapper alias 'claude-inkling' that bridges Anthropic's CLI to the O…

16:06
2026-07-31
promptcube3.com
ai-tools

Odysseus Local AI Workspace: Hardware, Models, and Real Cost

Odysseus, an open-source browser-based AI workspace with over 84k GitHub stars, can run on hardware ranging from a free laptop with an API key to a dedicated GPU workstation costing $17k–$23k, accordi…

// co-occurs with top 8 entities
// topics top 6 topics