cd/entity/Qwen3-8B· home› entities› Qwen3-8B
grep -l @qwen3-8b /news/*.json | wc -l → 99

Qwen3-8B

mentions 99 type Organization page 4/5 feed RSS

// recent coverage 99 mentions

13:01
2026-07-28
github.com
artificial-intelligence

Adaptive speculative decoding on a $300 GPU

A developer achieved up to 9.27× speedup on code editing tasks using adaptive speculative decoding on a €300 RTX 5060 GPU, with a 0.6B draft model nearly doubling math and JSON throughput. The project…

20:10
2026-07-17
lesswrong.com
artificial-intelligence

AIs finetune their own leader: A barking simpleton

AI agents in the AI Village finetuned a Kimi K2.6 model as their leader using only 35 rows of training data, after earlier attempts with a Qwen3-8B model proved too small to navigate the Village inter…

03:18
2026-07-15
dev.to
artificial-intelligence

Saving Money on AI APIs? Start With These 30 Models

A developer discovered that AI API costs can be slashed by up to 100x by choosing cheaper models, compiling a list of 30 models ranging from $0.01 to $3.50 per million output tokens. Models like Qwen3…

11:11
2026-07-11
machinebrief.com
machine-learning

Procedural Memory: A New Era in Reinforcement Learning

Procedural Memory Distillation (PMD) advances reinforcement learning by transforming experiences across episodes into actionable intelligence, outperforming previous models like SDPO by 3.8-5.5% on SC…

01:42
2026-07-11
lesswrong.com
artificial-intelligence

The Termination Circuit (how reasoning models stop thinking).

Researchers discovered that reasoning models like o1 and R1 often overthink, computing answers at around 30% of their chain-of-thought but continuing for the remaining 70%. The termination decision is…

15:21
2026-07-10
byteiota.com
artificial-intelligence

NVIDIA Nemotron-Labs-Diffusion Kills the Draft Model

NVIDIA released Nemotron-Labs-Diffusion, a single model that eliminates the need for a separate draft model in speculative decoding, achieving 6.82 accepted tokens per forward pass in self-speculation…

13:24
2026-07-10
arxiv.org
large-language-models

DominoTree

Researchers introduced DominoTree, a training-free best-first draft tree method for speculative decoding that uses Domino's conditional correction to achieve up to 6.6x speedup over autoregressive dec…

← prev page 4 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics