cd/entity/Qwen· home› entities› Qwen
grep -l @qwen /news/*.json | wc -l → 939

Qwen

mentions 939 type Organization page 9/47 feed RSS

// recent coverage 939 mentions

16:00
2026-09-03
blogs.nvidia.com
artificial-intelligence

Sparks Fly: NVIDIA Accelerates Local AI at IFA 2026

At IFA 2026, NVIDIA, Microsoft, and partners announced faster local AI inference and new tools for running agents locally on NVIDIA hardware, including up to 1.9x faster inference via llama.cpp and vL…

08:19
2026-09-03
github.com
large-language-models

Show HN: PicoLM v1.0-rc1

PicoLM v1.0-rc1, an LLM inference engine written in C99, has been released, supporting llama-2, GPT-2, Qwen 3.6/3.8(+MoE), and Gemma-3n models. The engine features CPU SIMD acceleration, CUDA/HIP supp…

04:13
2026-09-03
gist.github.com
machine-learning

README-omlx.md

A developer detailed their local AI setup using oMLX and Qwen models on an M5 Max Mac with 128GB unified memory, achieving 82-123 tok/s decode with a 35B-A3B MoE worker versus 15.5-17.8 tok/s for a 27…

20:45
2026-09-02
promptcube3.com
artificial-intelligence

Can we actually trust an LLM to build a secure codebase?

A new analysis questions whether large language models can be trusted to build secure codebases, highlighting the widening performance gap between frontier models and the persistent risks of AI-genera…

20:09
2026-09-02
byteiota.com
artificial-intelligence

Qwen3.8-Max-0902 Tops Coding Charts — Should You Switch?

Alibaba's Qwen team released Qwen3.8-Max-0902, a post-training update to its Qwen3.8-Max model, which debuted at #1 on Arena.ai's Code Arena: WebDev leaderboard with 1,691 points, edging Claude Opus 5…

17:00
2026-09-02
promptcube3.com
large-language-models

Qwen might actually leapfrog the 2.4T parameter giants

Qwen, DeepSeek, and GLM are closing the gap with trillion-parameter models by focusing on post-training optimization and extended reasoning rather than raw parameter count, according to an analysis on…

16:18
2026-09-02
fromtheterminal.substack.com
artificial-intelligence

The Model Is Not the Product Anymore

A developer argues that frontier model releases are no longer the industry's defining events, pointing to Hugging Face's State of Open Models: Summer 2026 report showing Chinese labs releasing far lar…

14:02
2026-09-02
github.com
large-language-models

WebLLM: high-performance in-browser LLM inference engine

MLC AI's WebLLM is a high-performance in-browser LLM inference engine that runs entirely in web browsers with WebGPU hardware acceleration, eliminating the need for server-side processing. It is fully…

06:34
2026-09-02
dev.to
large-language-models

LLM fine-tuning 101: a practical guide for developers

A developer's practical guide explains that fine-tuning large language models is now accessible to individual developers with consumer GPUs, thanks to techniques like LoRA and QLoRA. The guide details…

22:05
2026-09-01
dev.to
artificial-intelligence

The Edit That Fixed 4 Tasks and Broke 1

AgentSelfEdit, an open-source sidecar that rewrites its own system prompt from execution feedback, A/B tests prompt edits and promotes only statistically-proven winners. In a recent test, an LLM-propo…

20:04
2026-09-01
mkaz.blog
developer-tools

Pi, more than a coding harness

Developer mkaz built Halftrack, a personal trainer app powered by the Pi coding harness, which uses natural language commands to log and query running workouts. The app, available on GitHub, demonstra…

15:05
2026-09-01
dev.to
artificial-intelligence

Running Local AI Models on a Consumer GPU: A 2026 Field Test

A developer's field test of running local AI models on consumer GPUs in 2026 finds that memory capacity is the key constraint, with 8GB handling 3B-4B models and 32GB enabling 14B-30B models via quant…

← prev page 9 / 47 next →
// co-occurs with top 8 entities
// topics top 6 topics