cd/entity/Qwen3.8 27B· home› entities› Qwen3.8 27B
grep -l @qwen3.8 27b /news/*.json | wc -l → 23

Qwen3.8 27B

mentions 23 type Person page 1/2 feed RSS

// recent coverage 23 mentions

13:38
2026-09-29
forum.level1techs.com
ai-tools

Using Unsloth to finetune LLMs on Linux (Intel B70)

A user on the Level1Techs forum reported that Unsloth's default installation failed to run out of the box on an Intel Arc B70 GPU under Fedora 44, and was advised to manually compile the SYCL backend,…

14:00
2026-09-20
tokenstead.ai
large-language-models

Ternary Bonsai 2 27B

Ternary Bonsai 2 27B, a ternary-weight multimodal model built on a Qwen3.8 27B base, ships at roughly 6GB with weights restricted to {-1, 0, +1} at 1.76 bits per weight with FP16 group scales, accordi…

12:00
2026-09-18
julin.ai
large-language-models

Testing PrismML Bonsai 2

PrismML released Bonsai 2 27B, a ternary-weight compression of Qwen3.8 27B that cuts model size from about 56 GB to 5.9 GB while retaining roughly 98% of the original model's benchmark score. In a han…

22:34
2026-09-17
techcrunch.com
large-language-models

PrismML hopes its tiny LLM will change how we all use AI

PrismML released Bonsai 2 27B on Thursday, a compressed version of Alibaba's Qwen3.8 27B open-source model shrunk to 5.9 GB — a 9x to 10x memory reduction that matches 98% of Qwen's aggregate benchmar…

20:58
2026-09-17
runtimewire.com
large-language-models

PrismML launches a 5.9 GB Qwen3.8 model for local AI

PrismML launched Bonsai 2 27B on September 17th, an Apache 2.0-licensed ternary-weight version of Qwen3.8 27B whose language model fits in 5.95 GB and, according to the company, preserves 98.2% of the…

00:00
2026-09-17
digitalapplied.com
large-language-models

Run a 35B AI Model on a 24GB Mac: Two New Ways Compared

Edge0, an open-source framework posted to arXiv on September 16, 2026, reports running a 4-bit Qwen3.6-35B-A3B mixture-of-experts model at 20.4 tokens per second inside 2.9 GiB of peak active memory o…

00:00
2026-09-17
runagentrun.co.uk
ai-tools

Intel's free AI toolkit adds ten more models

Intel released OpenVINO 2026.4 on 16 September 2026, adding ten new model families for CPU and GPU inference, including Kokoro-82M text-to-speech, Qwen3-ASR speech-to-text, Muse Glimmer 30B, Qwen3.8 2…

20:09
2026-09-11
forum.level1techs.com
large-language-models

LLM Development System (GPUs Accounted for)

A user testing a multi-GPU LLM development machine reported that the system held up well under sustained load running the Qwen3.8 27B model, despite the machine shipping from 45 Drives with the wrong …

00:00
2026-09-08
mindstudio.ai
machine-learning

What Are GSQ and RCO? Das Lab's New LLM Quantization Method

Das Lab at the Institute of Science and Technology Austria (ISTA), the team behind GPTQ, introduced GSQ (Gumbel Softmax Quantization) and RCO (Riemannian Constrained Optimization), two linked techniqu…

00:00
2026-09-02
motherduck.com
large-language-models

Agents Don’t Query Like Humans Do

A new benchmark shows that the Qwen3.8 27B model outperforms GPT 5.6 Luna on the DABstep SQL benchmark while costing under 50 cents in electricity, over 17x less, according to a blog post by Alex Mona…

20:54
2026-08-31
promptcube3.com
large-language-models

Why Qwen3.

A hands-on benchmark of Qwen3.8 27B on an RTX 4090 found that 4-bit quantization (NF4 and AWQ INT4) preserves near-baseline quality, with MMLU scores of 59.8% and 59.5% versus 61.2% for FP16, while 1-…

13:41
2026-08-26
dev.to
large-language-models

I just killed hallucinations on my 2 bit Qwen3.8 27B

A developer has open-sourced SIMURG, a tool that monitors OpenAI-compatible endpoints and retries requests when it detects hallucinations. Testing with a 2-bit quantized Qwen3.8 27B model on an RTX 30…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics