cd/entity/GPU· home entities GPU
grep -l @gpu /news/*.json | wc -l → 110

GPU

mentions 110 type Organization page 4/6 feed RSS

// recent coverage 110 mentions

00:18
2026-06-19
dev.to
ai-agents

How I Run a 50-Agent AI Workforce on a Single 6GB GPU

A developer describes running ~50 local AI agents on a single 6GB GPU by using a lock-based queue, an eviction monitor, a resource governor, and a model router. The system serializes GPU access so onl…

00:00
2026-06-19
fergusfinn.com
ai-infrastructure

InfiniBand, RoCE, and all that

InfiniBand, a high-performance interconnect technology designed for Remote Direct Memory Access (RDMA), has become critical for AI training and inference workloads that require direct data movement be…

13:24
2026-06-17
letsdatascience.com
robotics

Nvidia showcases robots installing GPUs autonomously

Nvidia demonstrated an agentic robotics system called ENPIRE that taught fleets of robots to perform dexterous tasks, including handing a GPU to another arm and positioning it over a motherboard. The …

22:12
2026-06-16
0mean1sigma.com
machine-learning

2678x Faster Matrix Multiplication with a GPU

A developer achieved 2678x faster matrix multiplication using a GPU with CUDA, demonstrating how parallel processing on thousands of GPU cores reduces the O(N³) complexity of sequential matrix multipl…

18:57
2026-06-16
injuly.in
large-language-models

Inference cost at scale with napkin math

A technical analysis calculates the dollar cost per user for serving large language models at scale using napkin math, breaking down GPU resources, matrix multiplication costs, and attention mechanism…

20:40
2026-06-15
gilesthomas.com
machine-learning

Jax: Commitment Issues

JAX's default_device context manager places arrays on the specified device but does not commit them, allowing JAX to move them to other devices. This caused array lookups to take over a second by trig…

12:20
2026-06-15
dev.to
artificial-intelligence

The CPU Is Back in the Stack — and Nobody Budgeted for It

Agentic AI workloads are shifting the compute ratio, making CPUs the critical coordination substrate rather than a support component for GPUs. This inversion is exposed by current Xeon supply tightnes…

09:00
2026-06-15
infoworld.com
large-language-models

33 LLM metrics to watch closely

A comprehensive list of 33 metrics for evaluating large language models (LLMs) has been compiled, covering performance indicators such as time to first token, average tokens per second, throughput, er…

23:51
2026-06-14
neurophos.com
ai-chips

Neurophos OPU

Neurophos announced its Optical Processing Unit (OPU), a photonic AI chip achieving 0.47 ExaOPS at 300 TOPS/W, claiming 10,000x size reduction over traditional photonic tensor cores. The company says …

← prev page 4 / 6 next →
// co-occurs with top 8 entities
// topics top 6 topics