cd/entity/Qwen3· home› entities› Qwen3
grep -l @qwen3 /news/*.json | wc -l → 123

Qwen3

mentions 123 type Organization page 3/7 feed RSS

// recent coverage 123 mentions

15:54
2026-08-17
ziggit.dev
machine-learning

Fucina: a CPU-first tensor/autograd library in Zig

Matteo Grella released Fucina v0.1.0, a CPU-first tensor/autograd library written in pure Zig 0.16, featuring compile-time axis names in tensor types that make misaligned contractions a compile error.…

17:09
2026-08-16
sourcefeed.dev
artificial-intelligence

Fine-Tuning Can't Teach What Pretraining Never Saw

Researchers at the Max Planck Institute for Intelligent Systems, ELLIS Tübingen, and ETH Zürich built LittleLearner, a 5B-parameter language model pretrained exclusively on 88 billion tokens of U.S. e…

02:16
2026-08-14
hydradb.com
artificial-intelligence

Graph Engineering: Execution Graphs vs. Context Graphs

LangChain's blog post on graph engineering distinguishes execution graphs, which coordinate agent workflows, from context graphs, which represent shared domain knowledge, arguing that durable agent st…

09:39
2026-08-12
gist.github.com
large-language-models

Run a 4B LLM model locally on your mac

A developer has published a step-by-step guide for running a 4B parameter LLM locally on a Mac, using Mozilla's llamafile and a fine-tuned Qwen3 model from Hugging Face. The process involves downloadi…

16:23
2026-08-05
anicka.net
artificial-intelligence

Buddhist AI Research

Karma Electric, a Buddhist AI research group, reports that training language models to reason about suffering rather than memorize refusal patterns yields safety that persists even when compliance neu…

04:00
2026-08-04
machinebrief.com
artificial-intelligence

OoO-Spec: Out-of-Order Semantic Speculation for Fast Tool Calling

Researchers introduced OoO-Spec, an out-of-order semantic speculation method that speeds up LLM tool calling by predicting function choices and argument values in parallel. Across seven targets and th…

18:34
2026-08-03
twitter.com
artificial-intelligence

AI system Locus posttrains a model better than Qwen3

Locus, an automated AI research system developed by Sakana AI, achieves state-of-the-art results on PostTrainBench and surpasses the human post-trained Qwen3 1.7B model in the extended PostTrainBench+…

19:39
2026-07-30
lesswrong.com
large-language-models

Internal State Control is a General Property of LLMs

A replication study by the Second Look Fellowship finds that internal state control is a general property of large language models, with 14 models across the Qwen3, Gemma 3, and Tulu 3 families (0.3B …

← prev page 3 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics