cd/entity/RTX 4090· home entities RTX 4090
grep -l @rtx 4090 /news/*.json | wc -l → 30

RTX 4090

mentions 30 type Person page 1/2 feed RSS

// recent coverage 30 mentions

15:12
2026-07-27
openmodelmap.com
large-language-models

Kimi K3 Hardware Requirements

Kimi K3, the world's largest open-source model with 2.8 trillion parameters and a Mixture-of-Experts architecture requiring all 896 experts to be loaded into VRAM, needs a minimum of 6× H100 80GB GPUs…

19:46
2026-07-25
promptcube3.com
large-language-models

DeepSeek-R1 Local Deployment: My Hardware Struggles

A user reports that deploying the full 671B parameter DeepSeek-R1 model locally requires over 100GB of VRAM and is impractical on consumer hardware, with CUDA out-of-memory errors occurring even at sm…

14:50
2026-07-24
promptcube3.com
artificial-intelligence

Claude Code Workflow: Leveraging Open Weights for Local Dev

A developer reports that shifting to a hybrid AI workflow using Anthropic's Claude 3.5 Sonnet for architectural planning and a local Llama 3.1 8B model for unit test generation reduced token spend by …

04:00
2026-07-17
machinebrief.com
artificial-intelligence

Per-Token Fixed-Point Convergence in Depth-Recurrent Transformers

A 135M-class depth-recurrent transformer trained on FineWeb-Edu converges to a per-token fixed point, with mean successive-output KL divergence falling from 3.9e-1 at the second loop to 8.5e-6 by the …

07:39
2026-07-16
github.com
artificial-intelligence

Show HN: Trellis2.c – Local 3D generation with Vulkan and CUDA

Trellis2.c, a native executable for local 3D generation with Vulkan and CUDA backends, has been released on GitHub. The project aims to provide a lightweight alternative to Python/PyTorch runtimes, si…

06:22
2026-07-13
calcrecipe.com
artificial-intelligence

The Winners of the AI Era

The memory industry is emerging as a structural beneficiary of the AI era as autonomous AI agents and automation platforms drive demand for higher memory bandwidth, according to a Vault Track analysis…

12:31
2026-07-01
pub.towardsai.net
artificial-intelligence

I Tried Making Image Generation 90x Cheaper. Here’s What Worked.

A developer at a fashion-discovery startup reduced image generation costs by 90x by switching from Google's Gemini API to running Alibaba's open-source Qwen-Image-Edit model on an RTX 4090 GPU, levera…

01:11
2026-06-30
byteiota.com
artificial-intelligence

Ornith 1.0 Beats Claude at Coding — Runs on One GPU

DeepReinforce AI released Ornith 1.0, an open-source coding model family that scores 82.4 on SWE-Bench Verified, outperforming Claude Opus 4.7's 80.8, and runs the 35B variant locally on a single RTX …

13:11
2026-06-29
fergusfinn.com
machine-learning

What happens when you run a CUDA kernel?

NVIDIA's CUDA compiler pipeline transforms a simple vector addition kernel from PTX virtual assembly to SASS machine code through multiple compilation stages, including LLVM-based cicc and ptxas, befo…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics