cd/entity/AMD Instinct MI300X· home› entities› AMD Instinct MI300X
grep -l @amd instinct mi300x /news/*.json | wc -l → 14

AMD Instinct MI300X

mentions 14 type Person feed RSS

// recent coverage 14 mentions

18:09
2026-10-01
dev.to
ai-infrastructure

Evaluating Speculative Decoding in vLLM on AMD MI300X GPUs

A developer evaluated speculative decoding in vLLM on AMD Instinct MI300X and MI355X GPUs, testing draft methods including MTP, EAGLE-3, DFlash, and DSpark that propose candidate tokens verified in a …

00:00
2026-09-30
rocm.blogs.amd.com
ai-infrastructure

AIM: Unified User Experience from Profile Discovery to Deployment

AMD detailed its AMD Inference Microservices (AIMs), standardized Docker-based inference microservices that automatically select a runtime profile from inputs including model, precision, engine, laten…

21:36
2026-09-17
dev.to
ai-infrastructure

Serving Gemma 4 on an AMD MI300X: What $1.99 an Hour Buys

A developer published a step-by-step guide and open-source MCP toolkit for deploying Google's Gemma 4 E2B model to a single AMD Instinct MI300X GPU rented through AMD Developer Cloud at $1.99 per hour…

09:26
2026-09-07
vllm.ai
large-language-models

Speculative Decoding in vLLM on AMD GPUs

AMD's experiments with speculative decoding in vLLM on AMD Instinct MI300X and MI355X GPUs using the ROCm platform show that output-token throughput gains vary by drafting method, proposal length, mod…

14:31
2026-09-01
techstrong.ai
artificial-intelligence

PINNACLE Measures the Work, Not the Model

Signal65's new PINNACLE benchmark, developed with Kamiwaza, evaluates AI systems on whole enterprise task completion rather than model scores or token throughput, finding that 43 of 44 configurations …

00:00
2026-07-23
rocm.blogs.amd.com
ai-infrastructure

Onboard and Deploy Custom Models in AMD AI Workbench

AMD AI Workbench v2.0.0 now lets users onboard and deploy custom models from Hugging Face or private registries directly through its GUI, serving them as OpenAI-compatible endpoints via the same AMD I…

// co-occurs with top 8 entities
// topics top 6 topics