cd /news/ai-infrastructure/ibm-deploys-nvidia-hgx-b300-cluster-… · home topics ai-infrastructure article
[ARTICLE · art-91922] src=cryptobriefing.com ↗ pub= topic=ai-infrastructure verified=true sentiment=· neutral

IBM deploys Nvidia HGX B300 cluster on IBM Cloud, targeting regulated AI workloads

IBM has deployed Nvidia's HGX B300 GPU systems on IBM Cloud, offering up to 144 PFLOPS of FP4 compute per node for regulated industries such as financial services, healthcare, and government. The systems, available since April 29, integrate with Red Hat OpenShift and IBM's watsonx platform to support governed AI workflows, with serverless fleet access expected in May 2026.

read2 min views1 publishedAug 11, 2026
IBM deploys Nvidia HGX B300 cluster on IBM Cloud, targeting regulated AI workloads
Image: Cryptobriefing (auto-discovered)

Via nvidia.com

The deployment brings Blackwell Ultra GPUs capable of 144 petaflops to enterprises in financial services, healthcare, and government sectors

IBM has made Nvidia’s HGX B300 GPU systems available on IBM Cloud, giving regulated industries access to some of the most powerful AI training hardware on the market through a managed cloud environment. The systems went live around April 29, marking the latest step in an expanded partnership between the two companies that was first announced at GTC 2026 in March.

What’s inside the box #

Each HGX B300 system packs eight Blackwell Ultra GPUs into a single node. The raw compute numbers are staggering: up to 144 PFLOPS in FP4 precision and 72 PFLOPS in FP8. Each system offers more than 2 TB of shared HBM3e memory, which matters enormously for large language model training where the entire model needs to fit in GPU memory to avoid performance-killing data shuffling. The networking side runs at 800 Gb/s per GPU via Nvidia’s ConnectX-8 adapters, ensuring that multi-node training jobs don’t bottleneck at the interconnect layer.

The compliance angle #

The HGX B300 deployment integrates with Red Hat OpenShift and IBM’s watsonx platform, creating what IBM frames as a complete stack for governed AI workflows. OpenShift handles the container orchestration that lets workloads move between on-premises and cloud environments. Watsonx provides the AI development tooling and governance layer. Together, they allow enterprises to train and deploy models in a hybrid-cloud setup where sensitive data stays within approved boundaries.

Timing and rollout #

The initial availability targets select customers and regions, with a broader rollout expected to follow. IBM has also signaled that serverless fleet access for the HGX B300 systems is anticipated in May 2026, which would let customers consume GPU compute without managing underlying infrastructure.

The HGX B300 systems complement IBM Cloud’s existing H200 instances, which offer a lower performance tier for workloads that don’t require Blackwell Ultra-class hardware.

Disclosure: This article was edited by Editorial Team. For more information on how we create and review content, see our

Editorial Policy.

── more in #ai-infrastructure 4 stories · sorted by recency
── more on @ibm 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ibm-deploys-nvidia-h…] indexed:0 read:2min 2026-08-11 ·