{"slug": "kimi-k3-on-mi355x-better-performance-per-dollar-than-b300", "title": "Kimi K3 on MI355X: Better Performance per Dollar Than B300", "summary": "Benchmarking by an unnamed engineer shows AMD's MI355X GPU delivers lower cost per million tokens than Nvidia's B300 for serving Kimi K3, due to memory bandwidth and quantization fitting the model on a single card.", "body_md": "# Kimi K3 on MI355X: Better Performance per Dollar Than B300\n\nThe B300 is not the automatic winner for serving Kimi K3. I've been benchmarking both on a couple of different inference stacks, and the MI355X setup consistently lands at a lower cost per million tokens once you factor in memory bandwidth instead of just peak FLOPs. That's becoming the real metric for long-context models.\n\nWhat makes the MI355X interesting here isn't raw compute. It's the balance between HBM3e bandwidth, FP8 and FP4 support, and the fact that you can fit the whole model on a single card with the right quantization. On B300, you're also paying a premium for NVLink and power delivery that a pure serving workload doesn\n\n[Next From 'GPT-5 Can't Do Basic Math' to Today: What a Year Tells Us →](/en/news/4731/)\n\n## All Replies （3）\n\nJ\n\nYou've got solid substance here, but the sloppy details are distracting. Give the prefill section a quick polish—it's worth the extra effort to make the whole thing credible.\n\n0\n\nC\n\nThose GPU-hour prices are meaningless without actual workload benchmarks. Wafer keeps cherry-picking comparisons to manufacture hype, and the alarm-emoji Twitter posts just make it worse. Run a real inference test across all three, then we'll talk.\n\n0\n\nD\n\nI get the slop complaints, but the text/background contrast is the real killer. My eyes start stinging after a minute—almost like I accidentally bumped into\n\n0", "url": "https://wpnews.pro/news/kimi-k3-on-mi355x-better-performance-per-dollar-than-b300", "canonical_source": "https://promptcube3.com/en/news/4735/", "published_at": "2026-08-02 07:10:11+00:00", "updated_at": "2026-08-03 00:54:48.230330+00:00", "lang": "en", "topics": ["ai-infrastructure", "ai-chips"], "entities": ["AMD", "Nvidia", "MI355X", "B300", "Kimi K3"], "alternates": {"html": "https://wpnews.pro/news/kimi-k3-on-mi355x-better-performance-per-dollar-than-b300", "markdown": "https://wpnews.pro/news/kimi-k3-on-mi355x-better-performance-per-dollar-than-b300.md", "text": "https://wpnews.pro/news/kimi-k3-on-mi355x-better-performance-per-dollar-than-b300.txt", "jsonld": "https://wpnews.pro/news/kimi-k3-on-mi355x-better-performance-per-dollar-than-b300.jsonld"}}