cd/entity/RTX 5090· home› entities› RTX 5090
grep -l @rtx 5090 /news/*.json | wc -l → 112

RTX 5090

mentions 112 type Person page 1/6 feed RSS

// recent coverage 112 mentions

08:30
2026-10-07
discuss.huggingface.co
artificial-intelligence

Scaling Beatrix V3: a 376M byte-level model and the training

AbstractPhil released Beatrix V3, a 376-million-parameter byte-level language model with a 256-token vocabulary and 4,096-byte context, trained on 64.4 billion bytes across two RTX 5090 cards in 295 h…

04:44
2026-09-26
octet-stream.net
large-language-models

Getting Deeper into Local Inference

A personal LLM user reports that Qwen3.8-27B running locally via llama.cpp on an RTX 5090 laptop GPU with 24 GB of VRAM now handles Q&A, coding, sysadmin and web research without a cloud provider, del…

00:00
2026-09-23
mindstudio.ai
artificial-intelligence

Bonsai 2 27B: A 27B Model That Runs in Under 6GB on a Laptop

Prism ML released Bonsai 2 27B, a ternary-quantized version of Qwen3.8-27B that stores weights as -1, 0, or +1 and fits in roughly 6 to 8.6GB while retaining 98.2% of the original FP16 model's benchma…

18:09
2026-09-22
blog.comfy.org
ai-tools

Making the MiniMax H3 Video VAE 2x Faster

ComfyUI v0.36.0 ships optimizations that make the MiniMax H3 video VAE encode up to ~2.2x faster and decode ~1.4-2.7x faster, cutting a 1344x768, 129-frame encode-and-decode round trip from 24.3 to 12…

03:06
2026-09-22
gist.github.com
large-language-models

Qwen 3.8 27B on RTX 5090 at 90-120tps

A developer published a llama-server configuration that runs a Qwen 3.8 27B NVFP4 model with MTP speculative decoding on an RTX 5090, reporting throughput of 90-120 tokens per second. The setup uses a…

13:36
2026-09-19
tokenstead.ai
ai-chips

NVIDIA RTX PRO 5500 Blackwell 84GB

NVIDIA listed the RTX PRO 5500 Blackwell, an 84GB GDDR7 ECC workstation card with 21,760 CUDA cores, 1,398 GB/s of bandwidth, 600W power draw and PCIe Gen 5, as "Coming Soon" on nvidia.com in Septembe…

17:07
2026-09-08
openui.com
generative-ai

OUI-1: world's first model for Generative UI

Thesys Dev released OUI-1, the world's first model for Generative UI, a finetuned DiffusionGemma model that writes user interfaces in openui-lang and runs on consumer-grade GPUs like the RTX 5090 at F…

16:29
2026-09-05
richg42.blogspot.com
machine-learning

Neural block texture compression with CUDA

A developer built a neural block texture compression system that runs entirely on the GPU using CUDA 13.1, combining per-pixel selectors, per-4×4-block latents, and a tiny 1,767-weight MLP decoder tra…

page 1 / 6 next →
// co-occurs with top 8 entities
// topics top 6 topics