cd/entity/Unsloth· home› entities› Unsloth
grep -l @unsloth /news/*.json | wc -l → 90

Unsloth

mentions 90 type Organization page 3/5 feed RSS

// recent coverage 90 mentions

15:03
2026-08-14
unsloth.ai
artificial-intelligence

Unsloth - Qwen3.8 - How to Run Locally

Unsloth released dynamic GGUF quantizations for Qwen3.8, enabling the 27B model to run locally on 17-19GB VRAM setups and the 2.4T parameter model to run in 397GB via 1-bit quantization. The Qwen3.8 f…

17:00
2026-08-11
promptcube3.com
artificial-intelligence

Unsloth Desktop finally lets us train models locally without a

Unsloth Desktop, a new local AI training tool from Unsloth, enables users to train models locally without a cloud backend, supporting NVIDIA, AMD, Intel, and Mac hardware. It claims a 70% reduction in…

14:31
2026-08-11
unsloth.ai
ai-tools

Unsloth Desktop

Unsloth released Unsloth Desktop (Beta), a free, open-source app for running and training AI models locally on macOS, Windows, and Linux, supporting LLMs, diffusion image/video, MLX, GGUF, and audio m…

00:00
2026-08-11
mindstudio.ai
large-language-models

How to Run Nemotron 3.5 Lightning Locally on Your Own GPU

NVIDIA's open-weights Nemotron 3.5 Lightning mixture-of-experts model, with 30 billion total parameters but only 3 billion active, is designed for agentic grunt work and can be run locally on consumer…

21:42
2026-08-09
github.com
artificial-intelligence

AI Agent Qubitz

Qubitz, a local-first AI agent for GGUF models on llama.cpp, aims to make 7B–35B MCP/tool-capable LLMs more predictable and useful through a specialized harness and Agent Behavioral Contracts. It oper…

23:08
2026-08-05
thewatershed.markpesce.com
artificial-intelligence

Four Watersheds

The Watershed moment in AI is not a single event but four distinct watersheds at different scales and price points, according to a new analysis. The first, the Frontier Watershed, occurred in November…

16:08
2026-08-04
byteiota.com
artificial-intelligence

Qwen3.8-Max Is Open Weights: Switch Your API Today

Alibaba released Qwen3.8-Max, a 2.4-trillion-parameter sparse Mixture-of-Experts model, on August 3, claiming it outperforms GPT-5.6 Sol on key coding benchmarks and promising full open weights next w…

15:09
2026-08-04
sourcefeed.dev
machine-learning

An 8B Fine-Tune Now Fits in 4 GB of VRAM

Independent researcher Alpamys Makazhan released Soup, a Show HN project that fine-tunes a full Llama-3.1-8B model in NF4 quantization with a 3.32 GB VRAM peak at 119.6 tokens/sec on a 4 GB RTX 3050 L…

23:54
2026-08-03
github.com
artificial-intelligence

Unsloth: Run and Train Local LLMs

Unsloth, an open-source AI startup, released Unsloth Studio (Beta), a platform that lets users run and train text, audio, embedding, and vision models locally on Windows, Linux, and macOS, with suppor…

14:44
2026-08-03
tokenstead.ai
artificial-intelligence

DeepSeek V4 Flash 0731

DeepSeek released V4 Flash 0731, an iterative update of its Mixture-of-Experts model with 284B total parameters and 13B active per token, achieving an Artificial Analysis Intelligence Index of 50, up …

22:09
2026-08-02
byteiota.com
artificial-intelligence

GLM-5.2 Beats GPT-5.5 on SWE-bench — And You Can Self-Host It

Z.ai's GLM-5.2 scored 62.1 on SWE-bench Pro, beating GPT-5.5's 58.6, while costing $1.40 per million input tokens and $4.40 per million output tokens, making it 3.6x cheaper on input and 5.7x cheaper …

11:01
2026-07-30
promptcube3.com
large-language-models

Open-Weight Models Now Match Proprietary Titans

The accuracy gap between the best open-weight models and GPT-4o has shrunk to under 3% on structured data tasks, according to a developer's hands-on analysis. A fine-tuned Qwen2.5 72B model achieved 9…

22:33
2026-07-24
gilesthomas.com
large-language-models

Benchmarking Qwen 3.6 35B MoE (3B active) on an RTX 3090

Benchmarking Qwen 3.6 35B MoE (3B active) on an RTX 3090 using Unsloth's UD-IQ4_NL_XL quantisation achieved up to 140 tokens per second for generation and over 3,300 tok/s for prompt processing with a…

← prev page 3 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics