cd/entity/GB200· home› entities› GB200
grep -l @gb200 /news/*.json | wc -l → 20

GB200

mentions 20 type Organization feed RSS

// recent coverage 20 mentions

04:00
2026-10-03
arxiv.org
large-language-models

Fast Polynomial Transcendentals for LLMs

A new arXiv paper (2610.00049v1) reports that replacing native sigmoid, tanh, and SiLU activations with degree-3 or degree-4 bfloat16 polynomial programs improved complete training-step throughput on …

00:45
2026-08-26
servethehome.com
artificial-intelligence

OpenAI Jalapeno Custom AI ASIC at Hot Chips 2026

OpenAI unveiled Jalapeño, an in-house inference ASIC and system built with Broadcom, at Hot Chips 2026, claiming it outperforms NVIDIA GB200 and GB300 in throughput per kilowatt and latency for OpenAI…

17:37
2026-08-11
runtimewire.com
artificial-intelligence

Nvidia ships Nemotron 3.5 Lightning for single-GPU AI agents

Nvidia released Nemotron 3.5 Lightning on August 11, an open-weight reasoning model with 30 billion total parameters and 3 billion active per token, designed to run agent workloads on a single Nvidia …

17:22
2026-07-30
thediligencestack.com
ai-infrastructure

The Behind-the-Meter AI Buildout

Behind-the-meter (BTM) power is becoming a viable option for hyperscalers to accelerate AI infrastructure deployment, driven by a shift from training to inference workloads and improved rack-level pow…

15:01
2026-07-29
dwarkesh.com
artificial-intelligence

Why compute might get 10x+ more expensive in coming years

Compute costs could rise 10x or more in coming years, according to an analysis of AI lab economics. If a human-level software engineer could run on an H100 equivalent, that GPU should rent for over $2…

10:00
2026-07-24
fastcompany.com
ai-infrastructure

Where Nvidia is going, it doesn’t need cables

Nvidia's new Vera Rubin platform, named after the astronomer, eliminates nearly all cables from its AI server trays, reducing assembly time from two to three hours and cutting initial failure rates fr…

15:00
2026-07-21
theregister.com
ai-infrastructure

Nvidia shows off Vera Rubin platform for tokenmaxxing

Nvidia demonstrated its Vera Rubin platform at a lab tour in Sunnyvale, California, claiming the compute tray can be assembled in one minute versus 90 minutes for prior GB200 hardware, a 90x improveme…

15:21
2026-07-10
byteiota.com
artificial-intelligence

NVIDIA Nemotron-Labs-Diffusion Kills the Draft Model

NVIDIA released Nemotron-Labs-Diffusion, a single model that eliminates the need for a separate draft model in speculative decoding, achieving 6.82 accepted tokens per forward pass in self-speculation…

08:58
2026-06-13
narracomm.com
ai-infrastructure

Microsoft/OpenAI ecosystem continues massive NVIDIA GPU intake

Microsoft and OpenAI are planning to deploy hundreds of thousands of NVIDIA GB200 GPUs, alongside their custom Maia AI chips, as part of massive infrastructure expansion. Hyperscaler capital expenditu…

20:55
2026-05-20
xcancel.com
artificial-intelligence

Anthropic is expanding to Colossus2. Will use GB200

Anthropic announced it is expanding its partnership with SpaceX to increase its computing capacity at the Colossus 2 facility, utilizing GB200 hardware throughout June. The company will begin ramping …

// co-occurs with top 8 entities
// topics top 6 topics