SemiAnalysis reports that on AgentX (multi-turn, long-context agent workloads), a new AMD MI355x submission beats B300 on total tokens-per-dollar TCO at lower interactivity ranges, shouting out vLLM, AMD, and LMCache engineers—useful signal for agentic serving hardware choices.

Key Takeaways

  • Benchmark: AgentX targets multi-turn, long-context agent traffic.
  • Result: AMD MI355x beats B300 on tokens/$ TCO at lower interactivity.
  • Stack: shoutout to vLLM, AMD, and LMCache on agentic serving work.
ADSponsored