cd/entity/Qwen3.6-35B-A3B· home entities Qwen3.6-35B-A3B
grep -l @qwen3.6-35b-a3b /news/*.json | wc -l → 30

Qwen3.6-35B-A3B

mentions 30 type Organization page 1/2 feed RSS

// recent coverage 30 mentions

15:20
2026-08-25
github.com
artificial-intelligence

Show HN: I made a Raspberry with Qwen my local car AI

A developer built CarWatch, a Raspberry Pi 5-based in-car AI assistant that runs a 35B-parameter Qwen3.6-35B-A3B model locally at 3.5 tokens per second, achieving full offline functionality for voice,…

00:00
2026-08-24
mindstudio.ai
artificial-intelligence

Ornith-1.5-35B-A3B: A Self-Improving MoE Model for Coding Agents

Ornith AI released Ornith-1.5-35B-A3B, a 35B-parameter mixture-of-experts model that activates about 3B parameters per token, trained via a self-improvement loop that jointly optimizes task generation…

21:46
2026-08-21
github.com
artificial-intelligence

Run 290B+ frontier MoE models locally on your gaming PC

FlashML released FreeToken, an edge-native Mixture-of-Experts (MoE) serving engine that runs 290B+ parameter frontier MoE models locally on consumer gaming PCs at interactive speeds. The engine suppor…

00:00
2026-08-20
mindstudio.ai
artificial-intelligence

Ornith 1.5 35B-A3B Benchmarks: How It Stacks Up Against Qwen3.6

Deep Reinforce's Ornith 1.5 35B-A3B, a mixture-of-experts language model activating about 3 billion parameters per token, outperforms Qwen3.6-35B-A3B on every published coding and agentic benchmark, s…

13:15
2026-08-13
brentozar.com
artificial-intelligence

Testing AI with an Office Hours Question

Brent Ozar, a Microsoft SQL Server expert and podcast host, tested whether a fully local AI setup could answer an Office Hours question, using the MiniMax H3 image-to-video model and a local Qwen3.6-3…

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

How to Run fuse-1 Lite Locally: VRAM, Setup, and Formats

Fuse-1 Lite, a 5.72B parameter mixture-of-experts coding model from LiquidAI, can run locally with VRAM needs ranging from 3.36 GB in 4-bit quantized form to about 12 GB in full bfloat16 precision, ac…

06:11
2026-08-04
byteiota.com
artificial-intelligence

Swiftlet: Run an 80B LLM in 4.3 GB of RAM on Mac

Swiftlet, an open-source project from developer Leonickson, enables running an 80-billion-parameter Qwen3-Next-80B-A3B sparse Mixture-of-Experts model in just 4.3 GB of RAM on Apple Silicon Macs by ke…

05:09
2026-08-04
sourcefeed.dev
artificial-intelligence

Your SSD Is the New VRAM

Swiftlet, a 10,000-line Swift and Metal project under Apache 2.0, runs Alibaba's Qwen3-Next-80B-A3B at 4-bit in 4.3GB of peak RAM on a Mac, decoding at 4.5–5 tokens per second on an M5, and runs the Q…

12:08
2026-07-24
sourcefeed.dev
artificial-intelligence

Hetzner Is Quietly Commoditizing LLM Inference

Hetzner, the German cloud host known for low prices, has quietly launched a free, OpenAI-compatible LLM inference endpoint serving a single Qwen3.6-35B-A3B model, signaling that open-weight inference …

00:43
2026-07-22
jonready.com
artificial-intelligence

Agent swarms are great for local AI

Agent swarms make local AI rigs cost-effective for the first time, according to developer testing. A single-agent session on a 2×3090 rig costs $0.80 in API-equivalent tokens, while a swarm of 32 agen…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics