cd/entity/Unsloth· home› entities› Unsloth
grep -l @unsloth /news/*.json | wc -l → 90

Unsloth

mentions 90 type Organization page 2/5 feed RSS

// recent coverage 90 mentions

00:00
2026-08-28
mindstudio.ai
artificial-intelligence

GLM-5.3 Flash Hands-On: Multi-GPU Test, Coding, and Refusals

Z AI's GLM-5.3 Flash, tested via Unsloth's Dynamic 1-bit quantization (a roughly 93 GB file) across five GPUs using llama.cpp, generated simple outputs at 33-34 tokens per second but slowed to 14-15 t…

12:24
2026-08-27
unsloth.ai
large-language-models

Qwen3.8-Flash-Next: How to Run Locally

Qwen released Qwen3.8-Flash-Next, a 125B-parameter open-weight multimodal MoE model built on the Qwen4 architecture with a 262K context window, which outperforms Claude-4.6-Opus (Max) and can run loca…

22:11
2026-08-26
byteiota.com
large-language-models

GLM-5.3-Flash Open Weights Drop: Ox Alpha Identity Revealed

Zhipu AI revealed that the anonymous model "Ox Alpha" is GLM-5.3-Flash, an MIT-licensed open-weights model released on Hugging Face on August 26 after a six-day stealth run that accumulated 44 trillio…

15:40
2026-08-26
tokenstead.ai
artificial-intelligence

Qwen3.8-Flash-Next

Alibaba released Qwen3.8-Flash-Next, the first open-weight preview of the Qwen4 architecture, on 2026-08-26, featuring 125B total parameters with only 6B active per token and topping Qwen3.7-Plus-Base…

00:00
2026-08-25
digitalapplied.com
artificial-intelligence

M5 Ultra's 512GB: Can a Desktop Hold a Frontier Model?

Apple's M5 Ultra Mac Studio, announced August 25, 2026, offers up to 512GB of unified memory at 1.2TB/s, enough to hold trillion-parameter-class open-weight models like DeepSeek-V4-Pro (1.6T params, ~…

13:21
2026-08-21
teachmecoolstuff.com
large-language-models

Good Results when training Qwen 3 4B to learn a new domain

A developer successfully used Unsloth to perform continued pretraining on Qwen 3 4B, teaching the small local LLM to act as a travel advisor for a fictional city by targeting only 66 million parameter…

19:32
2026-08-20
news.ycombinator.com
large-language-models

Qwen3.8 Fetches Weird URLs

A developer using Qwen3.8-27B from Unsloth with dynamic 3.0 quants in opencode reported that the model attempted to fetch 16 weird URLs during a single session, including a URL with an Expires timesta…

09:16
2026-08-20
byteiota.com
artificial-intelligence

Unsloth Dynamic 3.0 GGUFs: Run Qwen3.8-27B on 17GB RAM

Unsloth released Dynamic 3.0 GGUFs, a quantization package that enables running the Qwen3.8-27B vision-reasoning model on 17GB of RAM, and it has been downloaded five million times in five days on Hug…

22:19
2026-08-19
tokenstead.ai
large-language-models

Qwen3.8-27B

Alibaba released the open-weight Qwen3.8-27B multimodal model on 2026-08-05 under Apache-2.0, featuring a 27B dense architecture with a vision encoder, 262,144-token native context, and hybrid thinkin…

18:36
2026-08-19
unsloth.ai
artificial-intelligence

Unsloth Dynamic 3.0 GGUFs

Unsloth released Dynamic v3.0 GGUFs for Qwen3.8-27B, claiming more than 10% better top-1% accuracy at the same size compared to every other provider. The new quants, which work with llama.cpp and Unsl…

16:08
2026-08-16
sourcefeed.dev
artificial-intelligence

Unsloth Turns Fine-Tuning Into a Desktop App

Unsloth, the open-source fine-tuning library, released Unsloth Desktop, a free beta app for macOS, Windows, and Linux that runs and trains LLMs, diffusion models, and audio models locally, directly co…

13:59
2026-08-16
github.com
artificial-intelligence

Rats!* (1994) source code reconstruction using Qwen 3.8

A developer reconstructed the source code for the 1994 Windows game Rats! using the local LLM Qwen3.8 27B BF16 on a 2024 MacBook Pro with an Apple M4 Max and 128 GB of memory, achieving an average rec…

19:08
2026-08-15
byteiota.com
artificial-intelligence

Meta Muse Glimmer 30B: Local AI Agent on One GPU

Meta Superintelligence Labs released Muse Glimmer on August 10, a 30B-parameter open-weight model under Apache 2.0, designed for local agentic workflows and tool calling. The model, available on Huggi…

07:12
2026-08-15
byteiota.com
artificial-intelligence

Qwen3.8-27B Is Out: The Local AI Model Developers Need

Alibaba released Qwen3.8-27B, a 27.78-billion-parameter open-weights multimodal model under Apache 2.0, with a 262,144-token native context window and configurable reasoning. Vendor-reported benchmark…

← prev page 2 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics