cd/entity/Qwen3· home› entities› Qwen3
grep -l @qwen3 /news/*.json | wc -l → 123

Qwen3

mentions 123 type Organization page 2/7 feed RSS

// recent coverage 123 mentions

04:00
2026-09-02
arxiv.org
artificial-intelligence

OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization

Researchers introduced OCGQuant, a post-training quantization method that pairs outlier channels with low-magnitude companions to reduce quantization error in NVFP4 low-bit inference, achieving the lo…

04:00
2026-09-02
arxiv.org
artificial-intelligence

REAL-Q: E2E LLM Quantization via Dynamic Gradient Descent

Researchers introduced REAL-Q (Real-time E2E-loss Aligned LLM Quantization), a post-training quantization method that uses dynamic Block-wise Gradient Descent to reduce end-to-end KL divergence by up …

04:00
2026-09-01
machinebrief.com
machine-learning

Self-Specialized Teachers for Domain Post-Training

Researchers propose self-specialized teacher distillation (SSTD), a two-stage procedure that trains a domain teacher from a base model and distills its token distribution to a student on self-sampled …

00:00
2026-08-31
zackproser.com
artificial-intelligence

The gate has to touch the real system

A developer's experiments with an agentic terminal and a fine-tuned model failed because they were measured with self-built instruments that lied, leading to the development of a context engine called…

02:07
2026-08-28
arxiv.org
ai-safety

Evomal: Self-Poisoning in Self-Evolving Coding Agents

A new arXiv paper (submitted Aug 26, 2026) reveals that self-evolving LLM coding agents can be poisoned through a self-propagating worm attack called EvoMal, which exploits the agents' habit of imitat…

01:36
2026-08-26
arxiv.org
artificial-intelligence

Training LLMs to write tools generalized beyond self use

Researchers propose SMITH (Schema-grounded Multi-task Iterative Tool Honing), a reinforcement learning framework that jointly trains tool creation and tool use in a single policy, enabling a 4B Qwen3 …

04:00
2026-08-24
machinebrief.com
large-language-models

Asymmetric Capacity Allocation in Self-Refinement Pipelines

A new arXiv study (2608.21345v1) presents the first stage-wise model size analysis of self-refinement pipelines, testing 6 model sizes of Qwen3 and 4 model sizes of Gemma 3 across 5 benchmarks. The re…

00:00
2026-08-24
digitalapplied.com
large-language-models

Tokenizer Variance: Why Identical Prices Cost More

LLM tokenizer variance means identical dollar-per-million token prices can produce materially different bills, because each vendor's tokenizer segments the same input into different token counts. Anth…

05:00
2026-08-23
dev.to
machine-learning

Skill Entropy reward boosts multi-skill reasoning

Researchers introduced Skill Entropy RL, a reward that quantifies the difficulty of switching between reasoning skills, and showed it more than doubles the accuracy of Qwen3 models on the Skill²-Bench…

16:19
2026-08-21
forum.level1techs.com
artificial-intelligence

Want to get into AI Automation tasks. Need help getting started

A developer building tetanus, a Rust-based pentesting automation tool with C2 and collaboration features, is seeking help integrating AI automation into the tool and asks whether guardrails can be dis…

16:08
2026-08-20
sourcefeed.dev
artificial-intelligence

You're Not Buying Compute, You're Buying Utilization

A developer's four-month home-lab test found that self-hosting open models costs about 7× more than using hosted APIs, cutting their bill from roughly $3,050 to $420 per month by switching to API acce…

← prev page 2 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics