cd/entity/H200· home entities H200
grep -l @h200 /news/*.json | wc -l → 62

H200

mentions 62 type Organization page 2/4 feed RSS

// recent coverage 62 mentions

13:46
2026-07-23
cast.ai
artificial-intelligence

GPU Cloud Pricing in 2026: What AI Compute Really Costs

AWS raised H200 GPU instance prices by 15% on January 4, 2026, the first GPU price increase in roughly two decades, while average GPU utilization across more than 23,000 Kubernetes clusters sits at ju…

05:09
2026-07-20
byteiota.com
artificial-intelligence

Kimi K3 Open Weights Drop July 27: The Developer Prep Guide

Moonshot AI's Kimi K3, a 2.8-trillion-parameter model that topped the Frontend Code Arena leaderboard on day one, releases its full open weights on July 27. The MXFP4 weights require approximately 1.4…

10:40
2026-07-17
cast.ai
artificial-intelligence

Fractional GPUs and GPU Rightsizing: Stop Wasting Whole Cards

Average GPU utilization in production Kubernetes clusters is just 5%, according to the Cast AI 2026 State of Kubernetes Optimization Report, meaning 95 cents of every GPU dollar goes to idle silicon. …

23:01
2026-07-16
pub.towardsai.net
large-language-models

Beyond the KV Cache: What Comes Next

A hardware analysis reveals that deploying 70-billion parameter models in FP16 requires moving 140 gigabytes of weights across the memory bus per token, creating a memory-bound bottleneck that limits …

09:43
2026-07-16
cast.ai
artificial-intelligence

Kubernetes GPU Autoscaling: Scale GPU Capacity to Real Demand

GPU utilization averages just 5% across production Kubernetes clusters, according to the Cast AI 2026 State of Kubernetes Optimization Report, meaning 95% of provisioned GPU capacity is idle at any mo…

21:12
2026-07-15
thinkingmachines.ai
artificial-intelligence

Inkling Model Card

Thinking Machines Lab, Inc. released Inkling, a general-purpose multimodal model with 975 billion total parameters and 41 billion active parameters, on July 15, 2026 under an Apache 2.0 license. The m…

14:16
2026-07-14
cryptobriefing.com
artificial-intelligence

Kalshi builds prediction markets for GPU computing power prices

Kalshi, the first federally regulated prediction market exchange in the US, has launched contracts allowing traders to speculate on the per-hour cost of running NVIDIA's H100, H200, B200, and RTX 5090…

03:47
2026-07-14
cryptobriefing.com
ai-chips

Nvidia halves Asia buyer list amid China chip crackdown

Nvidia is halving its list of approved customers across Asia and tightening distributor vetting as US export controls squeeze its China market share, which is projected to collapse from 66% in 2024 to…

← prev page 2 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics