Ax3.ai
Home / News / The Ax3 take
The RegisterAI1d ago

AMD MI355X shows 1.3× throughput vs H200 on select inference workloads

MLPerf inference results published this week show AMD's latest Instinct accelerator pulling ahead on LLM inference — though training remains an NVIDIA stronghold.

New MLPerf inference submissions show AMD's Instinct MI355X delivering up to 1.3× the throughput of NVIDIA's H200 on several large-language-model serving benchmarks, the strongest showing yet for AMD's accelerator line.

The gains concentrate in memory-bandwidth-bound inference at long context lengths, where the MI355X's HBM capacity advantage tells. Training submissions, by contrast, continue to favor NVIDIA's ecosystem, where software maturity and interconnect scale remain decisive.

Analysts noted that inference now represents the fastest-growing share of AI compute demand, which makes a credible second source commercially meaningful even without training parity.

The Ax3 take

What it means: For inference-heavy buyers, MI3XX capacity just became a legitimate lever on price — a credible alternative pressures H200 rates even if you never deploy it.

What to do about it: Quote both stacks. Ax3 brokers MI3XX and H200 capacity side by side; buyers running serving workloads should ask us to price the same requirement across both before committing.

More from the Pulse

Related coverage

ReutersNVIDIA2h ago

NVIDIA reveals Vera Rubin details — general availability slips to H2 2027

The Vera Rubin NVL72 platform, positioned as the successor to Blackwell, will ship six months later than previously indicated at GTC. Reservations open in Q3 2026.

Data Center DynamicsDatacenter6h ago

Bulgarian 800 MW campus enters phase-1 commissioning

The Sofia-region AI campus, backed by a consortium of Nordic hyperscalers, will bring 60 MW online in Q3 with a further 740 MW roadmap across 2027–2029.