cd/entity/Qwen 3.8 27B· home entities Qwen 3.8 27B
grep -l @qwen 3.8 27b /news/*.json | wc -l → 52

Qwen 3.8 27B

mentions 52 type Person page 2/3 feed RSS

// recent coverage 52 mentions

14:12
2026-08-26
forum.level1techs.com
large-language-models

Running qwen 3.6 / 3.8 on 3090+3080 over RPC?

A user running Qwen 3.8 27B on a 3090 reports that switching to ninfer, a Qwen-focused inference engine, doubled their token generation speed from 35-40 to 55-60 tokens per second. The user is conside…

13:32
2026-08-26
forum.level1techs.com
large-language-models

Make a custom qwen 3.8 27B abliterated for ninfer?

A user running ninfer's Qwen 3.8 27B model on an RTX 3090 is seeking guidance on creating an abliterated version of the model, noting that ninfer uses a custom file format that prevents simple configu…

22:32
2026-08-25
siliconangle.com
artificial-intelligence

Perplexity AI launches Portable Computer on-device AI agent

Perplexity AI Inc. launched Portable Computer, an on-device AI agent for desktops with Nvidia Corp. silicon, following reports that Nvidia is considering an investment valuing the startup at over $30 …

20:23
2026-08-25
discuss.huggingface.co
artificial-intelligence

Building Local: My 2026 Headless AI Server Journey

A developer reports that running Qwen 3.8 27B at Q5_K_M quantization on a dual AMD Radeon RX 7900 XT and 7800 XT setup achieves 20 tokens per second with a 256k context window, enabling autonomous mul…

13:00
2026-08-25
thedeepview.com
artificial-intelligence

Why Perplexity just launched a local agent with Nvidia

Perplexity launched Portable Computer, a local AI agent that runs open models on desktop hardware, with Nvidia's DGX Spark as a partner platform. The agent runs a post-trained version of Alibaba's Qwe…

00:00
2026-08-25
mindstudio.ai
artificial-intelligence

DeepSeek V4 Flash on One RTX 3090: Real Tokens-Per-Second Numbers

DeepSeek V4 Flash, a mixture-of-experts model, ran at roughly 10 to 11 tokens per second on a single RTX 3090 with 192GB of system RAM in tests by FreeToken's desktop app, while a dense Qwen 3.8 27B m…

14:42
2026-08-23
gladlabs.io
artificial-intelligence

Qwen 3.8 27B Needs 22,000 Tokens to Draw a Pelican

Alibaba's Qwen 3.8 27B open-weight vision model, released under Apache 2.0, takes 21 minutes and 22,276 reasoning tokens to draw a pelican riding a bicycle as an SVG, according to Simon Willison's tes…

03:27
2026-08-23
forum.level1techs.com
large-language-models

Running qwen 3.6 / 2.8 on 3090+3080 over RPC?

A user reports running Qwen 3 Coder 30B A3B, Qwen 3.6 27B, and Qwen 3.8 27B on a local machine with a 7800X3D, 64GB DDR5, and an RTX 3090 24GB, achieving about 70 tokens per second on Qwen 3.8 27B, an…

22:10
2026-08-22
forum.level1techs.com
artificial-intelligence

What i run with my Strix Halo

A user reports running large language models on a 128GB Bosgame M5 Strix Halo mini PC purchased used for 1800€ on eBay, achieving 30 tokens per second decode with Qwen 3.8 27B and 50 tokens per second…

19:44
2026-08-22
forum.level1techs.com
large-language-models

Local one-shots, you say? Yep, we're there with Qwen 3.8 27B

Qwen 3.8 27B, a local LLM from Alibaba's Qwen team, achieves milestone-level coding performance on consumer hardware, with a user reporting it as the first local model that convinced them to integrate…

07:21
2026-08-22
github.com
generative-ai

Show HN: MiniMax H3 on a 16GB Mac, 5 days after open weights

VPIPE, a new macOS app from developer tgo-app-dev, runs MiniMax H3, a 33B multimodal model generating video and audio, on Apple Silicon Macs with as little as 16 GB RAM, achieving a 5-second 0.5 MP 24…

00:00
2026-08-21
runagentrun.co.uk
artificial-intelligence

Qwen 3.8 27B holds up at small sizes

A quality study from Kingy AI, a local-AI benchmarking blog, found that Qwen 3.8 27B, a 27-billion-parameter vision model, achieves roughly 92% top-token agreement with the full-precision reference at…

07:48
2026-08-20
williamcallahan.com
large-language-models

Qwen 3.8 27B is a great open model

Alibaba Cloud's Qwen 3.8 27B open-weights model delivers roughly double the intelligence and capability of last year's OpenAI gpt-oss models on the same hardware, according to software engineer Willia…

← prev page 2 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics