cd/entity/RTX 3090· home entities RTX 3090
grep -l @rtx 3090 /news/*.json | wc -l → 57

RTX 3090

mentions 57 type Person page 2/3 feed RSS

// recent coverage 57 mentions

14:37
2026-07-20
dev.to
large-language-models

Bonsai-27B on a Single 3090: What Works and What Doesn't

A developer tested the Bonsai-27B model on a single RTX 3090 using a Q4_K_M GGUF quant from prism-ml, achieving ~28 tok/s at 4K context and ~19 tok/s at 16K. The model excelled at structured extractio…

12:01
2026-07-14
pub.towardsai.net
artificial-intelligence

The Open Source Counter-Strike: Running Local Coding Agents

Open-source coding agents can now run entirely on local consumer hardware such as a 2020-era NVIDIA RTX 3090, eliminating usage-based pricing from AI vendors. The setup uses a quantized Gemma4 model s…

00:23
2026-07-09
gilesthomas.com
artificial-intelligence

poppy the training box, part 1: the beginnings

A developer repurposed an old small-form-factor PC named 'poppy' into a dedicated machine for local LLM training, upgrading its case and power supply to accommodate future multi-GPU setups. The projec…

10:32
2026-07-04
dev.to
large-language-models

Solving the GPU Pinning Saga and Gemma's Meta-Commentary

Glad Labs fixed a GPU pinning issue where LiteLLM 1.89.2's global api_base override prevented per-model routing, causing vision tasks to cold-load onto the wrong GPU. The team also hardened content gu…

15:03
2026-07-03
github.com
large-language-models

Jamesob's guide to running SOTA LLMs locally

Jamesob published a guide on building a local system to run state-of-the-art large language models, detailing hardware configurations ranging from $2k to $40k. The setup uses multiple RTX GPUs and PCI…

11:14
2026-06-30
byteiota.com
large-language-models

Qwen3.6 MTP in llama.cpp: 27B Model Now 1.7x Faster

On May 16, 2026, llama.cpp merged Multi-Token Prediction (MTP) support, enabling 1.7x to 2.4x faster local inference for Qwen3.6 27B models with no accuracy loss or extra downloads. The MTP head is em…

← prev page 2 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics