cd/entity/AMD Instinct MI355X· home› entities› AMD Instinct MI355X
grep -l @amd instinct mi355x /news/*.json | wc -l → 15

AMD Instinct MI355X

mentions 15 type Person feed RSS

// recent coverage 15 mentions

18:09
2026-10-01
dev.to
ai-infrastructure

Evaluating Speculative Decoding in vLLM on AMD MI300X GPUs

A developer evaluated speculative decoding in vLLM on AMD Instinct MI300X and MI355X GPUs, testing draft methods including MTP, EAGLE-3, DFlash, and DSpark that propose candidate tokens verified in a …

00:00
2026-09-17
rocm.blogs.amd.com
machine-learning

Technical Dive into AMD MLPerf Inference v6.1 Submission

AMD reported leadership scores and expanded benchmark coverage in MLPerf Inference v6.1, whose results were released on September 16, 2026. AMD and its partners submitted validated results across dlrm…

00:00
2026-09-17
rocm.blogs.amd.com
ai-research

Reproducing AMD MLPerf Inference v6.1 Submission Results

AMD published a step-by-step recipe for reproducing its MLPerf Inference v6.1 submission results on AMD Instinct MI355X, MI350X, and MI350P GPUs, marking the company's fifth consecutive round of MLPer…

09:26
2026-09-07
vllm.ai
large-language-models

Speculative Decoding in vLLM on AMD GPUs

AMD's experiments with speculative decoding in vLLM on AMD Instinct MI300X and MI355X GPUs using the ROCm platform show that output-token throughput gains vary by drafting method, proposal length, mod…

00:00
2026-08-24
rocm.blogs.amd.com
artificial-intelligence

Serving 64Mi-Token Contexts on One AMD Instinct™ MI355X Node

On a single 8-GPU AMD Instinct MI355X node, AMD served Kimi Linear 48B-A3B with context lengths from 1024 tokens to 64Mi (67,108,864), using vLLM with FP8 KV cache and tensor parallelism 8, recording …

00:00
2026-07-23
rocm.blogs.amd.com
artificial-intelligence

Serve Kimi-K2.5-MXFP4 on MI355X with ATOM

AMD shows how to serve the pre-quantized amd/Kimi-K2.5-MXFP4 checkpoint on AMD Instinct MI355X GPUs using ATOM, a lightweight vLLM-like framework that integrates AITER kernels and exposes an OpenAI-co…

// co-occurs with top 8 entities
// topics top 6 topics