cd/entity/Qwen· home› entities› Qwen
grep -l @qwen /news/*.json | wc -l → 944

Qwen

mentions 944 type Organization page 39/48 feed RSS

// recent coverage 944 mentions

11:29
2026-06-25
discuss.huggingface.co
large-language-models

LLM "curving" via prompting

A researcher has developed a prompting technique called 'LLM curving' that shifts large language models from token-by-token prediction to a holistic self-organization mode, aiming to improve reasoning…

10:01
2026-06-25
discuss.huggingface.co
large-language-models

Deepseek? Qwen?

A single H200 GPU with 141GB HBM3e cannot comfortably run DeepSeek V4 Flash (284B total, 13B active parameters) due to VRAM constraints, even with 2TB system RAM for offloading. The model requires an …

09:50
2026-06-25
oracomputing.com
large-language-models

ORA: Smaller Models. Same Intelligence

Ora Computing launched an automated LLM compression engine that reduces model size by up to 70% with minimal accuracy loss, enabling deployment on edge devices, on-prem servers, or cloud infrastructur…

08:05
2026-06-25
dev.to
artificial-intelligence

I Replaced 2.5 Hours of Daily Busywork with a $0 AI Agent Setup

A developer replaced 2.5 hours of daily busywork with a $0 AI agent setup running on a Mac Mini M4. The system uses local LLMs (Ollama with Qwen models), Python scripts, and cron jobs to automate emai…

04:03
2026-06-25
devclubhouse.com
artificial-intelligence

The distillation attack no API can fully block

Anthropic accused Alibaba of executing the largest known distillation attack against its Claude AI model, involving 28.8 million queries from 25,000 fraudulent accounts between April and June 2025. Th…

01:36
2026-06-25
pangram.com
artificial-intelligence

Exploring the internal representations of Pangram 3.3.2

Pangram Labs researchers explored the internal representations of their AI detection model Pangram 3.3.2 using document-level analysis of activations across layers, aiming to understand what the model…

22:48
2026-06-24
letsdatascience.com
ai-safety

Anthropic Accuses Alibaba of Distilling Claude

Anthropic accused Alibaba's Qwen AI lab of conducting the largest distillation attack on its Claude model, using nearly 25,000 fraudulent accounts to make 29 million exchanges between April 22 and Jun…

22:08
2026-06-24
dev.to
developer-tools

Cli-Modelarium 0.1.4: 10 LLM providers now, with Qwen and GLM

Cli-Modelarium 0.1.4 adds support for Alibaba's Qwen models (via DashScope) and Z.AI's GLM models, bringing the total to 10 cloud LLM providers. The command-line tool enables side-by-side comparison o…

15:19
2026-06-24
dev.to
large-language-models

GLM 5.2: Reasoning Effort Is the Cost Lever

Zhipu's GLM 5.2 open-weight model, available on Synthorai at roughly one-sixth of frontier per-token prices, achieves frontier-level benchmarks but its per-task cost varies by over an order of magnitu…

12:03
2026-06-24
devclubhouse.com
large-language-models

Simulating the World Inside the LLM

Alibaba's Qwen team released Qwen-AgentWorld, a language model that simulates complex environments natively, replacing external simulators for training AI agents. The model, trained on over 10 million…

← prev page 39 / 48 next →
// co-occurs with top 8 entities
// topics top 6 topics