cd/entity/DeepSpeed· home entities DeepSpeed
grep -l @deepspeed /news/*.json | wc -l → 12

DeepSpeed

mentions 12 type Organization feed RSS

// recent coverage 12 mentions

15:09
2026-08-04
sourcefeed.dev
machine-learning

An 8B Fine-Tune Now Fits in 4 GB of VRAM

Independent researcher Alpamys Makazhan released Soup, a Show HN project that fine-tunes a full Llama-3.1-8B model in NF4 quantization with a 3.32 GB VRAM peak at 119.6 tokens/sec on a 4 GB RTX 3050 L…

16:50
2026-07-04
github.com
large-language-models

OpenScience: Workbench for scientific research using custom LLMs

Synthetic Sciences launched OpenScience, an open-source AI workbench that automates the full scientific research loop—literature review, hypothesis formation, code writing, experiment execution, and w…

15:05
2026-06-03
pytorch.org
machine-learning

Using Muon Optimizer with DeepSpeed

DeepSpeed has integrated the Muon Optimizer, a memory-efficient optimizer that uses a single momentum buffer and Newton-Schulz orthogonalization to improve training convergence, particularly for 2D we…

15:25
2026-04-29
pytorch.org
large-language-models

Introducing AutoSP

Researchers at Microsoft have introduced AutoSP, a compiler-based solution that automatically converts standard training code into multi-GPU sequence parallel code for long-context language model trai…

// co-occurs with top 8 entities
// topics top 6 topics