cd/entity/Qwen· home entities Qwen
grep -l @qwen /news/*.json | wc -l → 686

Qwen

mentions 686 type Organization page 29/35 feed RSS

// recent coverage 686 mentions

23:26
2026-06-17
discuss.huggingface.co
large-language-models

Introducing KerasFormers: "Transformers" for Keras 3!

KerasFormers, an open-source library built entirely in Keras 3, launches with over 100 transformer models spanning vision, language, multimodal, and speech, supporting seamless execution on TensorFlow…

20:29
2026-06-17
news.ycombinator.com
large-language-models

Ask HN: What models are powering your product's AI features?

A Hacker News user asked the community which AI models power their products' daily features, revealing that many teams choose models based on cost and latency rather than rigorous evaluation. The user…

20:15
2026-06-17
github.com
large-language-models

Show HN: Selora – local model for Home Assistant

Selora AI Local, an open-source Qwen-based model for Home Assistant, has been released. The model features four specialized LoRA adapters for answers, clarifications, automations, and commands, runnin…

18:30
2026-06-17
castform.com
large-language-models

I post-trained a model to reliably roll a die

A developer post-trained a language model to roll a die, revealing that models default to outputting 4 due to training data bias. Standard reinforcement learning fails to encourage exploration, but ad…

17:58
2026-06-17
lesswrong.com
ai-safety

Porting MACHIAVELLI To Inspect

A developer ported the MACHIAVELLI benchmark, which measures unethical AI agent behavior, to the Inspect evaluation framework to make it easier for evaluators to use. The re-implementation is now offi…

13:09
2026-06-17
dev.to
large-language-models

DeepSeek or Qwen 3 Max? I Ran Both for a Month on Client Work

A freelance developer tested DeepSeek V4 Flash, DeepSeek V4 Pro, Qwen3-32B, GLM-4 Plus, and GPT-4o for a legal document classification system over a month. DeepSeek V4 Flash achieved comparable accura…

08:13
2026-06-17
sebastianraschka.com
artificial-intelligence

VibeThinker-3B and the Strength of Post-Training

WeiboAI released VibeThinker-3B, a 3.09B-parameter coding and reasoning model built on Qwen2.5-Coder-3B that achieves performance close to much larger systems through extensive post-training, includin…

04:00
2026-06-17
arxiv.org
artificial-intelligence

Dissecting model behavior through agent trajectories

Researchers introduced the Simple Strands Agent (SSA) to minimize the 'intent-execution gap' between AI models and their harnesses, improving agent performance. Analyzing 138k trajectories, they found…

00:00
2026-06-17
runagentrun.co.uk
ai-infrastructure

OpenRouter fans prompts to match Claude Fable 5

OpenRouter launched Fusion, a routing layer that sends a single prompt to multiple AI models in parallel and synthesizes their outputs, achieving performance comparable to Anthropic's Claude Fable 5 a…

20:21
2026-06-16
dev.to
artificial-intelligence

Notion AI's Pricing Trap: Why I Went Open Source Instead

A developer abandoned Notion AI after its pricing ballooned, opting for open-source alternatives. Benchmarking showed Notion AI's optimized 2026 stack offered 40-65% cost reduction but relied on commu…

20:00
2026-06-16
loomcycle.dev
ai-agents

loomcycle 1.0 is here. Substrate complete. What's next.

Loomcycle 1.0, a feature-complete agentic runtime, is now available under Apache-2.0 license. The Go-based binary supports six LLM providers, 19 built-in tools, multi-replica high availability, and pa…

19:35
2026-06-16
dev.to
artificial-intelligence

Airtable AI From Scratch: A Freelance Dev's Cost Breakdown

A freelance developer rebuilt their AI stack around Airtable AI, reducing monthly API costs from $89 to $14—an 84% drop—by switching from GPT-4o to cheaper models like DeepSeek V4 Flash and Qwen3-32B …

15:33
2026-06-16
mmlac.com
artificial-intelligence

Local AI Is Not Ready for Coding. Yet?

A test of local AI models for coding tasks found that consumer hardware cannot run models capable of autonomous software development. Devstral 24B and Qwen3.5 122B-a10b failed to complete a simple fil…

← prev page 29 / 35 next →
// co-occurs with top 8 entities
// topics top 6 topics