cd/entity/MI355X· home entities MI355X
grep -l @mi355x /news/*.json | wc -l → 24

MI355X

mentions 24 type Organization page 1/2 feed RSS

// recent coverage 24 mentions

18:12
2026-08-28
tokenstead.ai
artificial-intelligence

AMD ROCm 10: what the launch means for local AI

AMD shipped ROCm 10 on August 26, 2026, a ground-up rebuild of its GPU software stack that unifies support for Instinct, Radeon, and Ryzen AI on Linux and Windows, with AMD reporting an average 3.3x i…

00:45
2026-08-26
servethehome.com
artificial-intelligence

OpenAI Jalapeno Custom AI ASIC at Hot Chips 2026

OpenAI unveiled Jalapeño, an in-house inference ASIC and system built with Broadcom, at Hot Chips 2026, claiming it outperforms NVIDIA GB200 and GB300 in throughput per kilowatt and latency for OpenAI…

07:10
2026-08-02
promptcube3.com
ai-infrastructure

Kimi K3 on MI355X: Better Performance per Dollar Than B300

Benchmarking by an unnamed engineer shows AMD's MI355X GPU delivers lower cost per million tokens than Nvidia's B300 for serving Kimi K3, due to memory bandwidth and quantization fitting the model on …

06:03
2026-08-02
promptcube3.com
artificial-intelligence

AI News Digest · Aug 2

AI News Digest for August 2 covers 29 stories, including Kimi K3 on MI355X offering better performance per dollar than B300, Figure AI's F.03 ladder climb raising autonomy or OSHA concerns, and Google…

11:54
2026-08-01
runinfra.ai
ai-infrastructure

$0.09 and $290.12 are both the price of 1M output tokens

A cost analysis across 24 providers and 378 GPU rental rates found that the price of one million output tokens ranges from $0.09 on a single AMD MI355X to $290.12 on eight NVIDIA H100s, with the gap d…

16:48
2026-07-23
sdxcentral.com
artificial-intelligence

AMD launches Instinct MI400 Series GPUs for AI workloads

AMD launched the Instinct MI400 Series GPUs for AI workloads, including the MI455X for frontier AI and the MI430X for sovereign AI and HPC. Built on CDNA 5 architecture and a 2nm process, the MI455X f…

00:00
2026-07-08
rocm.blogs.amd.com
artificial-intelligence

SGLang-ATOM: Bring ROCm-Native Acceleration to SGLang Serving

AMD introduced SGLang-ATOM, a bridge connecting the SGLang serving framework with ATOM's ROCm-native execution path to accelerate large language model inference on AMD Instinct GPUs. The integration u…

00:00
2026-07-06
epics.tech
large-language-models

The Tool Breaks as the Model Arrives

Open-weight model GLM-5.2 matched Claude Opus at a fraction of the cost, global LLM token spending fell 20% from its May peak, and a GPT-5.5 Codex flaw caused quiet output degradation, signaling a mar…

22:07
2026-06-30
bargo.ai
ai-infrastructure

GPU Compute Tightness Index

The GPU Compute Tightness Index fell to 43.8/100 as of June 29, 2026, entering Loose territory after a monthly drop of 13.7 points, indicating abundant idle AI infrastructure and supply exceeding near…

00:00
2026-06-24
rocm.blogs.amd.com
machine-learning

DP Attention and TBO for DeepSeek-V4 on MI355X

AMD introduces DP Attention and Two-Batch Overlap (TBO) optimizations for DeepSeek-V4 inference on MI355X GPUs, using a coordinated prefill scheduler called PrefillDelayer to reduce padding waste and …

16:05
2026-06-17
fortran-lang.discourse.group
artificial-intelligence

The AI era is pulling FP64 hardware away from scientific HPC

The AI boom is pulling GPU vendors away from double-precision (FP64) hardware essential for scientific HPC, as NVIDIA, AMD, and Intel prioritize low-precision AI cores. New chips like NVIDIA's B200 an…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics