cd/entity/Qwen 3.8 27B· home entities Qwen 3.8 27B
grep -l @qwen 3.8 27b /news/*.json | wc -l → 52

Qwen 3.8 27B

mentions 52 type Person page 1/3 feed RSS

// recent coverage 52 mentions

02:04
2026-09-18
byteshape.com
ai-tools

Shapelearn Qwen 3.8 27B (13.1 GB VRAM)

ByteShape released full ShapeLearn GGUF quantizations for Qwen 3.8 27B, with all five models sitting on the measured quality-speed frontier across six GPU comparisons, the company reported. The releas…

12:58
2026-09-16
twitter.com
large-language-models

Operating System powered by Qwen 3.8 27B at 1950 tokens/SEC

A developer built a minimal Python web server that turns Cerebras-hosted inference of Alibaba's Qwen 3.8 27B model into a live operating system with zero apps on disk, streaming tokens at 1,950 tokens…

20:00
2026-09-15
akitaonrails.com
large-language-models

Novo LLM Benchmark v4: retestando 39 LLMs (Parte 2)

AkitaOnRails published Part 2 of its LLM Benchmark v4, retesting 39 large language models on a seven-sprint Rails app suite seeded with 14 real CVE-based sabotages, after the author rejected the entir…

19:02
2026-09-15
forum.level1techs.com
ai-infrastructure

Homelab critique

A homelab user identified as gessha detailed a home setup that uses two Nvidia RTX 3090 GPUs in an i5-8600K workstation to host Qwen 3.8 27B and Qwen 3-VL 4B models via llama-server, alongside three m…

00:52
2026-09-15
forum.level1techs.com
ai-infrastructure

Trying hybrid inference (GPU+CPU) with 200-400B models

A user running hybrid GPU+CPU inference on 4x Nvidia V100 GPUs reported beating an Nvidia RTX 5090 in token generation (TG) for Qwen 3.8 27B and outperforming two DGX Spark systems with Qwen 3.8 Flash…

20:08
2026-09-07
github.com
large-language-models

Peer-to-peer LLM inference in browser tabs, Qwen 3.8 27B

SwarmLLM, an open-source project by developer Nehanth, enables a Qwen 3.8 27B model to run across browser tabs on multiple devices, with each device holding a slice of the model and passing 10 KB acti…

11:15
2026-09-07
byteiota.com
artificial-intelligence

Qwen 3.8 27B on Cerebras: 1,500 Tokens Per Second

Cerebras added Alibaba's Qwen 3.8 27B to its public API, delivering 1,500 tokens per second, enabling a 300-word response in under half a second. The model scores 61.7% on SWE-bench Pro, ahead of Clau…

20:05
2026-09-04
getreadyforagents.com
artificial-intelligence

Mistral announces agentic search with data opt-out guarantees

Mistral announced Agentic Search, a feature enabling autonomous search workflows with a data opt-out guarantee for training. The company also highlighted inference optimization for agentic use cases, …

13:34
2026-09-04
thenewway.ai
artificial-intelligence

OpenAI ships GPT-6 Astra, and it costs 2.5x GPT-5.6 Sol

OpenAI shipped GPT-6 Astra on Thursday at $10 per million input tokens and $50 per million output, 2.5 times the price of GPT-5.6 Sol, and rated it 'Critical' for cyber capability under its Preparedne…

08:24
2026-09-04
snipvote.com
artificial-intelligence

Qwen 3.8 27B available on Cerebras at 1500 tokens/s

Cerebras Systems now offers Qwen 3.8 27B at 1500 tokens per second on its wafer-scale hardware, reducing end-to-end latency to approximately 18 milliseconds per token and enabling real-time multi-step…

18:32
2026-09-03
inference-docs.cerebras.ai
artificial-intelligence

Qwen 3.8 27B available on Cerebras at 1500 tok/SEC

Cerebras Systems announced that the Qwen 3.8 27B model is now available on its platform at 1500 tokens per second, with all public models served unpruned and using selective weight-only quantization f…

13:50
2026-08-29
motherduck.com
artificial-intelligence

Agentic SQL for Free: Qwen3.8 27B and DuckDB

Qwen 3.8 27B, running locally on a 16GB RAM MacBook Pro, outperformed OpenAI's GPT 5.6 Luna Max on the DABstep benchmark at over 17 times lower cost, with electricity costs under $0.50 versus over $8.…

00:00
2026-08-28
motherduck.com
artificial-intelligence

Agentic SQL for Free with Qwen3.8 27B and DuckDB

Qwen 3.8 27B, an open-weight model, outperformed OpenAI's GPT 5.6 Luna Max on the DABstep benchmark while running locally on a laptop, costing under $0.50 in electricity versus over $8 for Luna Max, a…

19:50
2026-08-26
techstrong.ai
artificial-intelligence

Perplexity’s Portable Computer Keeps More Agent Work Local

Perplexity launched Portable Computer, a local version of its agentic AI software that runs on-device to keep data off the cloud by default, available initially to Pro and Max subscribers on Nvidia's …

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics