cd/entity/InfiniBand· home entities InfiniBand
grep -l @infiniband /news/*.json | wc -l → 19

InfiniBand

mentions 19 type Organization feed RSS

// recent coverage 19 mentions

11:11
2026-08-24
247wallst.com
artificial-intelligence

The Bear Case on Agents and Compute is Wrong So I Buy More Nvidia

NVIDIA Corporation (NASDAQ: NVDA) reported Q1 FY27 Data Center revenue of $75.246 billion, up 92% year over year, and total revenue of $81.61 billion, up 85.23%, as CEO Jensen Huang said 'Agentic AI h…

05:16
2026-08-15
promptcube3.com
ai-infrastructure

Nvidia is chasing a 500 billion dollar target that has Wall

Nvidia is pursuing a $500 billion market valuation target by integrating its InfiniBand networking, CUDA software, and Blackwell architecture into a proprietary full-stack ecosystem, creating high swi…

04:00
2026-08-03
arxiv.org
artificial-intelligence

Topology-Aware Data Movement for Disaggregated GPU Inference

A new arXiv paper (arXiv:2607.28633v1) proposes a topology-aware transfer orchestrator for disaggregated GPU inference, claiming existing systems like DistServe, Splitwise, and Mooncake ignore that ba…

15:00
2026-07-23
blogs.cisco.com
artificial-intelligence

Why AI inference is becoming a networking issue

Cisco's new white paper, 'A Day in the Life of a Prompt,' warns that AI inference is becoming a networking issue as data movement emerges as the primary bottleneck for GPU performance. The paper forec…

00:00
2026-07-21
fergusfinn.com
artificial-intelligence

NVLink, NVSwitch, and all that

NVIDIA's NVLink and NVSwitch technologies form a scale-up fabric that connects GPUs tightly enough to behave as a single machine, contrasting with scale-out fabrics like InfiniBand and RoCE that link …

17:08
2026-07-14
aleksagordic.com
artificial-intelligence

TPU and GPU Clusters: The Anatomy of Collective Communication

Google's TPU clusters use 2D or 3D torus topologies with 4 or 6 nearest neighbors per chip, while NVIDIA GPU clusters rely on fat-tree topologies over InfiniBand, according to a July 14, 2026 technica…

22:50
2026-06-26
dev.to
artificial-intelligence

Why AI Clusters Fail Even When GPUs Are Idle

AI clusters often underperform despite powerful GPUs because the GPUs are idle due to bottlenecks in data loading, CPU preprocessing, network communication, or storage contention. A developer explains…

00:00
2026-06-19
fergusfinn.com
ai-infrastructure

InfiniBand, RoCE, and all that

InfiniBand, a high-performance interconnect technology designed for Remote Direct Memory Access (RDMA), has become critical for AI training and inference workloads that require direct data movement be…

19:52
2026-06-11
developer.nvidia.com
ai-infrastructure

One-Click Multi-Tenant Security with  NVIDIA Quantum InfiniBand

NVIDIA introduced intent-based security profiles in its Unified Fabric Manager for Quantum InfiniBand, enabling network administrators to configure multi-tenant fabric security with a single click. Th…

15:05
2026-06-09
ubuntu.com
ai-infrastructure

What is RDMA over Converged Ethernet (RoCE)?

Canonical published an explainer on RDMA over Converged Ethernet (RoCE), a technology that runs the RDMA programming model over standard Ethernet networks to bypass the kernel and reduce CPU involveme…

12:03
2026-06-02
ubuntu.com
ai-infrastructure

What is InfiniBand?

InfiniBand is a purpose-built network interconnect designed to link compute, storage, and accelerator nodes with high bandwidth and low, predictable latency. The technology integrates Remote Direct Me…

01:03
2026-05-29
blog.hellas.ai
ai-infrastructure

Thunderbolt-ibverbs: We have InfiniBand at home

A developer created a Linux kernel module and userspace shim that transforms a standard USB4 connection into a low-latency InfiniBand device, enabling distributed AI inference across two 128GB Strix H…

08:30
2026-05-28
mistral.ai
ai-infrastructure

Mistral Compute? I hear Mistral Cloud

Mistral has launched Mistral Cloud, a private, integrated compute stack featuring GB300 GPUs, 1:1 InfiniBand XDR fabric, and SLURM plus Kubernetes orchestration. The platform offers tiered deployment …

// co-occurs with top 8 entities
// topics top 6 topics