cd/entity/Unsloth· home entities Unsloth
grep -l @unsloth /news/*.json | wc -l → 63

Unsloth

mentions 63 type Organization page 2/4 feed RSS

// recent coverage 63 mentions

16:08
2026-08-04
byteiota.com
artificial-intelligence

Qwen3.8-Max Is Open Weights: Switch Your API Today

Alibaba released Qwen3.8-Max, a 2.4-trillion-parameter sparse Mixture-of-Experts model, on August 3, claiming it outperforms GPT-5.6 Sol on key coding benchmarks and promising full open weights next w…

15:09
2026-08-04
sourcefeed.dev
machine-learning

An 8B Fine-Tune Now Fits in 4 GB of VRAM

Independent researcher Alpamys Makazhan released Soup, a Show HN project that fine-tunes a full Llama-3.1-8B model in NF4 quantization with a 3.32 GB VRAM peak at 119.6 tokens/sec on a 4 GB RTX 3050 L…

23:54
2026-08-03
github.com
artificial-intelligence

Unsloth: Run and Train Local LLMs

Unsloth, an open-source AI startup, released Unsloth Studio (Beta), a platform that lets users run and train text, audio, embedding, and vision models locally on Windows, Linux, and macOS, with suppor…

14:44
2026-08-03
tokenstead.ai
artificial-intelligence

DeepSeek V4 Flash 0731

DeepSeek released V4 Flash 0731, an iterative update of its Mixture-of-Experts model with 284B total parameters and 13B active per token, achieving an Artificial Analysis Intelligence Index of 50, up …

22:09
2026-08-02
byteiota.com
artificial-intelligence

GLM-5.2 Beats GPT-5.5 on SWE-bench — And You Can Self-Host It

Z.ai's GLM-5.2 scored 62.1 on SWE-bench Pro, beating GPT-5.5's 58.6, while costing $1.40 per million input tokens and $4.40 per million output tokens, making it 3.6x cheaper on input and 5.7x cheaper …

11:01
2026-07-30
promptcube3.com
large-language-models

Open-Weight Models Now Match Proprietary Titans

The accuracy gap between the best open-weight models and GPT-4o has shrunk to under 3% on structured data tasks, according to a developer's hands-on analysis. A fine-tuned Qwen2.5 72B model achieved 9…

22:33
2026-07-24
gilesthomas.com
large-language-models

Benchmarking Qwen 3.6 35B MoE (3B active) on an RTX 3090

Benchmarking Qwen 3.6 35B MoE (3B active) on an RTX 3090 using Unsloth's UD-IQ4_NL_XL quantisation achieved up to 140 tokens per second for generation and over 3,300 tok/s for prompt processing with a…

19:28
2026-07-20
github.com
artificial-intelligence

Unsloth: Introducing AMD support

Unsloth announced AMD GPU support for local LLM training and inference, enabling 500+ models to run up to 2× faster with 70% less VRAM on AMD Radeon, Instinct, Ryzen, and data center GPUs. The release…

15:01
2026-07-14
pub.towardsai.net
large-language-models

How I Fine-Tuned an 8B AI Model to Reason on a Free GPU

A student fine-tuned Meta's Llama 3.1 8B model for multi-step mathematical reasoning using Unsloth, LoRA, and a 'Silent Coder' approach, all within the RAM limits of a free Google Colab instance with …

13:14
2026-07-11
huggingface.co
artificial-intelligence

Introduction to Reinforcement Learning and Its Role in LLMs

A new course chapter introduces reinforcement learning (RL) and its application to training large language models (LLMs), explaining core concepts such as agent, environment, action, reward, and polic…

22:05
2026-07-10
dev.to
large-language-models

Despligue local GLM 5.2

A developer detailed the local deployment of GLM 5.2, a 753B MoE model with 40B active parameters requiring 1.51TB of RAM in BF16 precision. Unsloth's dynamic quantizations, such as UD-Q3_K_XL at 343G…

15:26
2026-07-10
aws.amazon.com
artificial-intelligence

Deploying quantized models on Amazon SageMaker AI with Unsloth

Amazon Web Services announced a partnership with Unsloth to deploy quantized foundation models on Amazon SageMaker AI, reducing memory usage and serving costs while maintaining accuracy. The collabora…

← prev page 2 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics