cd/entity/H100· home entities H100
grep -l @h100 /news/*.json | wc -l → 150

H100

mentions 150 type Organization page 7/8 feed RSS

// recent coverage 150 mentions

14:30
2026-06-12
dev.to
machine-learning

nvidia-smi Reports 97% Utilization While the GPU Sits Idle

A developer found that `nvidia-smi` reported 97% GPU utilization on an H100 cluster while actual training throughput was less than half of expected benchmarks. Tracing via eBPF revealed the GPU was id…

17:00
2026-06-10
pytorch.org
large-language-models

Portable vLLM Model Inference Kernels in Helion

Helion kernels were integrated into vLLM for FP8 inference using Qwen3 models and evaluated across NVIDIA H100 and B200 GPUs. The experiments demonstrated that Helion provides a productive PyTorch-nat…

09:56
2026-06-06
equixly.com
artificial-intelligence

We fight GPU scarcity without compromise

Equixly engineers report that GPU scarcity is a structural problem driven by hyperscaler stockpiling and rising energy costs, not a temporary supply-chain issue. The company argues that standard load-…

15:09
2026-06-03
dev.to
ai-products

Replicate vs deAPI: Price Comparison for AI Inference (2026)

A developer compared the costs of AI inference on Replicate and deAPI across four common tasks, finding that deAPI's task-based billing is often cheaper than Replicate's time-based or per-unit pricing…

17:52
2026-06-02
fergusfinn.com
ai-infrastructure

Bringing Up DeepSeek-V4-Flash on AMD MI300X

AMD's MI300X accelerator, with 192GB of HBM3 memory and roughly half the list price of NVIDIA's H100, remains underutilized due to software incompatibilities. As of early May 2026, running vLLM with D…

23:18
2026-05-30
categoryvc.com
ai-chips

AI Hardware

Modern GPUs spend most of their time during AI inference waiting for data, as memory bandwidth cannot keep pace with compute throughput. This fundamental bottleneck has driven the AI hardware market, …

← prev page 7 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics