cd/entity/Qwen· home entities Qwen
grep -l @qwen /news/*.json | wc -l → 686

Qwen

mentions 686 type Organization page 32/35 feed RSS

// recent coverage 686 mentions

15:24
2026-06-05
dev.to
large-language-models

Friday Fixes: Housekeeping the Homelab and Hub

A developer updated a homelab's local LLM stack, catching up llama.cpp by 469 builds and upgrading the Qwen generation model from 3.5 to 3.6, the embedding model from nomic v1.5 to v2-moe, and adding …

04:00
2026-06-05
arxiv.org
large-language-models

LoRi: Low-Rank Distillation for Implicit Reasoning

Researchers have developed LoRi, a low-rank distillation framework that improves implicit reasoning in large language models by aligning teacher and student reasoning trajectories within a shared low-…

02:44
2026-06-05
arxiv.org
machine-learning

OPRD: On-Policy Representation Distillation

Researchers have introduced On-Policy Representation Distillation (OPRD), a method that aligns student and teacher model representations across selected layers during training, bypassing the language …

19:33
2026-06-04
news.ycombinator.com
large-language-models

Words do not have determined meanings

A new demo called SRT-Introspect reveals how large language models like Qwen fix word meanings during generation, contradicting the reflexive, approximated nature of human language. The tool surfaces …

00:29
2026-06-04
github.com
large-language-models

TensorSharp: Open-Source Local LLM Inference Engine

TensorSharp, a new open-source C# inference engine, now enables developers to run large language models locally using GGUF files. The engine supports multiple model architectures including Gemma 4, Qw…

07:01
2026-06-03
dev.to
artificial-intelligence

Self-hosted video creation is coming

A developer is migrating video creation to an entirely self-hosted, open-source model, aiming to eliminate external API dependencies. The project currently uses only two APIs—11labs for voice and Vert…

14:13
2026-06-02
huggingface.co
ai-agents

Holo3.1: Fast & Local Computer Use Agents

Holo3.1, a new family of computer-use agents, is now available with improved robustness across web, desktop, and mobile environments. The release introduces quantized checkpoints for local inference, …

00:00
2026-06-02
tomtunguz.com
artificial-intelligence

The Thriving Ecosystem of Open Models

Open-weight models now account for 69.1% of named token volume on the OpenRouter API platform, compared to 30.9% for closed models, according to the latest platform data. The shift, driven by rapid co…

12:32
2026-05-30
maltebuettner.eu
large-language-models

DocumentAI Visual Benchmark - GPT 5.5, Gemini 3.5, Qwen...

A new benchmark evaluating DocumentAI models on bounding box accuracy shows GPT-5.5 and Gemini 3.5 leading with 67.7% and 67.5% scores respectively, while Qwen, Kimi, and Mistral trail significantly. …

20:42
2026-05-29
old.reddit.com
large-language-models

Claude Opus 4.8 may have distilled Qwen

An error page from Reddit blocked access to a story about Claude Opus 4.8 potentially having distilled Qwen. The request was denied due to a network policy, preventing the retrieval of the article's c…

18:13
2026-05-29
gist.github.com
large-language-models

Benchmark Qwen3.6 27B on Modal

A developer benchmarked the Qwen3.6 27B model on Modal using llama.cpp, deploying a serverless pipeline that downloads GGUF shards from Hugging Face and runs perplexity evaluation on an A100-80GB GPU.…

07:00
2026-05-29
github.com
ai-tools

SharkBay – a local macOS workbench for coding-agent CLIs

SharkBay, a new macOS workbench for managing multiple AI coding agents, launched as an open-source tool that lets developers run Claude Code, Codex, Gemini, and other agents from a single workspace. T…

06:01
2026-05-29
thedeepview.com
artificial-intelligence

AI safety benchmark reveals deeper LLM weaknesses

TELUS Digital released a comprehensive AI safety benchmark using 620,000 attack simulations across 34 models from 10 global AI labs, revealing that some models engaged with harmful requests more than …

18:38
2026-05-28
robertkarl.net
large-language-models

Qwen vs. Proust: Injecting novels into a local model's prompt

A developer injected the full text of Marcel Proust's *Swann's Way* into the prompt of a heavily quantized Qwen 9B 3.5 model to test where local language models break under noise during coding tasks. …

← prev page 32 / 35 next →
// co-occurs with top 8 entities
// topics top 6 topics