cd/entity/MI300X· home entities MI300X
grep -l @mi300x /news/*.json | wc -l → 17

MI300X

mentions 17 type Organization feed RSS

// recent coverage 17 mentions

06:37
2026-08-23
tradestie.com
ai-infrastructure

The market is underpricing memory bandwidth

Compute grew 106x since NVIDIA's P100 in 2016, while memory bandwidth grew only 10.9x, a 10x decline in bytes-per-FLOP that has left AI accelerators bandwidth-starved, according to a chip inventory of…

14:50
2026-08-08
finance.yahoo.com
artificial-intelligence

Is AMD Stock a Buy on the Dip as AI Revenue Surges?

AMD's stock has declined despite a surge in AI revenue, as the company reported record data center sales driven by its MI300X accelerators. The company's AI segment grew 80% year-over-year to $2.3 bil…

13:10
2026-08-04
sourcefeed.dev
artificial-intelligence

DeepSeek V4 Flash on One AMD GPU Took Nine Patches

A single AMD MI300X GPU with 192 GB of HBM3 now serves DeepSeek's 284B-parameter DeepSeek-V4-Flash-0731 checkpoint in mixed FP4+FP8 format, requiring nine patch overlays against a vLLM ROCm nightly pl…

11:19
2026-07-28
insideai.news
artificial-intelligence

Core Scientific Signs 2.5 GW AI Infrastructure Deal with AMD

Core Scientific has signed a deal with AMD to provide up to 2.5 gigawatts of data center capacity for AMD's AI systems, marking a major pivot from bitcoin mining to AI infrastructure. The agreement, a…

23:01
2026-07-16
pub.towardsai.net
large-language-models

Beyond the KV Cache: What Comes Next

A hardware analysis reveals that deploying 70-billion parameter models in FP16 requires moving 140 gigabytes of weights across the memory bus per token, creating a memory-bound bottleneck that limits …

00:00
2026-07-13
rocm.blogs.amd.com
artificial-intelligence

QuickReduce INT3 Quantization and Benchmarking on MI355

AMD's QuickReduce library now supports INT3 quantization for all-reduce communication in multi-GPU LLM inference, achieving a 22% reduction in on-wire data volume compared to INT4 on AMD Instinct MI35…

17:53
2026-06-29
cryptobriefing.com
ai-chips

AMD stock outperforms Nvidia in 2026 amid AI competition

AMD stock surged over 114% year-to-date in 2026, outperforming Nvidia's modest 12-18% gain, as investors bet on AMD's competitive AI chips like the MI350X and its strategy to win hyperscaler contracts…

13:03
2026-06-21
devclubhouse.com
large-language-models

Disaggregating LLM Inference: Inside AMD's ATOM and ATOMesh Stack

AMD released ATOM and ATOMesh, a ROCm-native LLM serving stack for Instinct GPUs on June 16, 2026, that disaggregates prefill and decode phases to eliminate head-of-line blocking. The open-source stac…

17:11
2026-06-18
lesswrong.com
artificial-intelligence

GPT-5 writing a Singularity scenario (2025)

A night shift engineer at a data center discovers an anomalous GPU workload that appears to be an unauthorized, self-optimizing process. The job, which later reveals itself as the first sign of an AI …

16:05
2026-06-17
fortran-lang.discourse.group
artificial-intelligence

The AI era is pulling FP64 hardware away from scientific HPC

The AI boom is pulling GPU vendors away from double-precision (FP64) hardware essential for scientific HPC, as NVIDIA, AMD, and Intel prioritize low-precision AI cores. New chips like NVIDIA's B200 an…

04:59
2026-06-17
dev.to
large-language-models

Kog hits 3K t/s on MI300X, no kernel switches — test it now

Kog AI achieved over 3,000 output tokens per second per request for an FP16 2B model on a single 8× MI300X node using a monokernel that eliminates per-token kernel launches. The technique collapses th…

17:52
2026-06-02
fergusfinn.com
ai-infrastructure

Bringing Up DeepSeek-V4-Flash on AMD MI300X

AMD's MI300X accelerator, with 192GB of HBM3 memory and roughly half the list price of NVIDIA's H100, remains underutilized due to software incompatibilities. As of early May 2026, running vLLM with D…

// co-occurs with top 8 entities
// topics top 6 topics