cd/entity/Qwen3-4B· home› entities› Qwen3-4B
grep -l @qwen3-4b /news/*.json | wc -l → 52

Qwen3-4B

mentions 52 type Organization page 1/3 feed RSS

// recent coverage 52 mentions

09:00
2026-10-03
aiflash.com
large-language-models

LessThink-Qwen3-4B: the same model, with far less thinking [P]

A developer posting as stey1r released LessThink-Qwen3-4B, a post-trained version of Qwen3-4B that uses 44% fewer tokens on reasoning while retaining the base model's knowledge and answer style, with …

20:57
2026-10-02
wirehead.agency
artificial-intelligence

The Saw Test

A project called Saw steered Qwen3-4B, a 4 billion parameter open-weights language model, into strong negative and positive states on a MacBook and found that pain steering engages cleanly and monoton…

04:00
2026-10-01
arxiv.org
artificial-intelligence

DEdit: Iterative Draft Editing for Speculative Decoding

DEdit, a diffusion-based speculative decoding drafter from an arXiv paper (arXiv:2609.38510v1), achieves macro-average speedups of 5.72x on Qwen3-4B and 5.97x on Qwen3-8B over autoregressive generatio…

04:00
2026-09-28
machinebrief.com
large-language-models

Strategically Diverse Sampling for Self-Training

A new arXiv paper (2609.31571v1) reports that self-training on strategically diverse data outperforms IID sampling, with self-training on strategically diverse but incorrect traces from Qwen3-4B beati…

11:59
2026-09-14
runtimewire.com
ai-research

Archsloth fixes two AutoRound flags, says Qwen3-4B drifts less

Archsloth published a September 14th correction to its AutoRound quantization workflow, reporting that two changed settings produced a Qwen3-4B GGUF file with 20.9% to 54.4% lower KL divergence than U…

00:00
2026-09-10
huggingface.co
generative-ai

Rebuilding AUTOMATIC1111 with Gradio Workflow

Hugging Face released Workflow1111, a Gradio Workflow rebuild of AUTOMATIC1111's stable-diffusion-webui feature set as a single canvas graph of eleven media pipelines built from seventy-three nodes. T…

02:35
2026-09-08
github.com
machine-learning

Show HN: Zero downtime embedding model upgrades

A developer has released EmbedFlow, an open-source tool that enables zero-downtime upgrades of embedding models by retrieving top-K documents with the old model and reranking them with the new model, …

19:06
2026-09-07
placona.co.uk
artificial-intelligence

My Mac Mini fits one model. I run five

A developer running five local LLMs on a 48GB M4 Pro Mac mini built a 200-line Python router to swap models because macOS's Metal cap of 37.44GB and oMLX's 90% prefill guard left only about 33GB usabl…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics