cd/entity/InfiniBand· home› entities› InfiniBand
grep -l @infiniband /news/*.json | wc -l → 23

InfiniBand

mentions 23 type Organization page 1/2 feed RSS

// recent coverage 23 mentions

11:00
2026-09-10
dev.to
large-language-models

Training a 3.8B LLM to 0.384 CORE for $998!

A developer trained a 3.8-billion-parameter dense language model to a CORE score of 0.384 for a total budget of $998, according to an account of the Little LM project. The run relied on Grouped Query …

11:00
2026-08-30
promptcube3.com
ai-infrastructure

Nvidia is winning the AI war by selling entire ecosystems

Nvidia is winning the AI war by selling entire ecosystems rather than just chips, leveraging CUDA software gravity, TensorRT, NeMo, and Mellanox-owned InfiniBand networking to lock developers into its…

17:45
2026-08-29
promptcube3.com
ai-infrastructure

Nvidia is winning the AI race by fixing data center bottlenecks

Nvidia is winning the AI race by focusing on data center interconnect and networking technologies rather than just raw chip performance, according to an analysis. The company's NVLink, Mellanox-based …

11:11
2026-08-24
247wallst.com
artificial-intelligence

The Bear Case on Agents and Compute is Wrong So I Buy More Nvidia

NVIDIA Corporation (NASDAQ: NVDA) reported Q1 FY27 Data Center revenue of $75.246 billion, up 92% year over year, and total revenue of $81.61 billion, up 85.23%, as CEO Jensen Huang said 'Agentic AI h…

05:16
2026-08-15
promptcube3.com
ai-infrastructure

Nvidia is chasing a 500 billion dollar target that has Wall

Nvidia is pursuing a $500 billion market valuation target by integrating its InfiniBand networking, CUDA software, and Blackwell architecture into a proprietary full-stack ecosystem, creating high swi…

04:00
2026-08-03
arxiv.org
artificial-intelligence

Topology-Aware Data Movement for Disaggregated GPU Inference

A new arXiv paper (arXiv:2607.28633v1) proposes a topology-aware transfer orchestrator for disaggregated GPU inference, claiming existing systems like DistServe, Splitwise, and Mooncake ignore that ba…

15:00
2026-07-23
blogs.cisco.com
artificial-intelligence

Why AI inference is becoming a networking issue

Cisco's new white paper, 'A Day in the Life of a Prompt,' warns that AI inference is becoming a networking issue as data movement emerges as the primary bottleneck for GPU performance. The paper forec…

00:00
2026-07-21
fergusfinn.com
artificial-intelligence

NVLink, NVSwitch, and all that

NVIDIA's NVLink and NVSwitch technologies form a scale-up fabric that connects GPUs tightly enough to behave as a single machine, contrasting with scale-out fabrics like InfiniBand and RoCE that link …

17:08
2026-07-14
aleksagordic.com
artificial-intelligence

TPU and GPU Clusters: The Anatomy of Collective Communication

Google's TPU clusters use 2D or 3D torus topologies with 4 or 6 nearest neighbors per chip, while NVIDIA GPU clusters rely on fat-tree topologies over InfiniBand, according to a July 14, 2026 technica…

22:50
2026-06-26
dev.to
artificial-intelligence

Why AI Clusters Fail Even When GPUs Are Idle

AI clusters often underperform despite powerful GPUs because the GPUs are idle due to bottlenecks in data loading, CPU preprocessing, network communication, or storage contention. A developer explains…

00:00
2026-06-19
fergusfinn.com
ai-infrastructure

InfiniBand, RoCE, and all that

InfiniBand, a high-performance interconnect technology designed for Remote Direct Memory Access (RDMA), has become critical for AI training and inference workloads that require direct data movement be…

19:52
2026-06-11
developer.nvidia.com
ai-infrastructure

One-Click Multi-Tenant Security with  NVIDIA Quantum InfiniBand

NVIDIA introduced intent-based security profiles in its Unified Fabric Manager for Quantum InfiniBand, enabling network administrators to configure multi-tenant fabric security with a single click. Th…

15:05
2026-06-09
ubuntu.com
ai-infrastructure

What is RDMA over Converged Ethernet (RoCE)?

Canonical published an explainer on RDMA over Converged Ethernet (RoCE), a technology that runs the RDMA programming model over standard Ethernet networks to bypass the kernel and reduce CPU involveme…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics