PUBLICATIONSPAPER RECORD2026

ProtocolBench: Which LLM MultiAgent Protocol to Choose?

Hongyi Du, Jiaqi Su, Jisen Li, Lijie Ding, Yingxuan Yang, Peixuan Han, Xiangru Tang, Kunlun Zhu, Jiaxuan You

ICML 2026First authorarXiv:2510.17149

Abstract

ORIGINAL TEXT

As large-scale multi-agent systems evolve, the communication protocol layer has become a critical yet under-evaluated factor shaping performance and reliability. Despite the existence of diverse protocols (A2A, ACP, ANP, Agora, etc.), the selection of them is often intuition-driven and lacks standardized guidance. We introduce ProtocolBench, a benchmark that systematically compares agent protocols along four measurable axes: task success, end-to-end latency, message or byte overhead, and robustness under failures. On ProtocolBench, the choice of protocol significantly influences system behavior. In the Streaming Queue scenario, overall completion time varies by up to 36.5% across protocols, and mean end-to-end latency differs by 3.48 s. Under Fail-Storm Recovery, resilience also differs consistently across protocols. Beyond evaluation, we present ProtocolRouter, a learnable protocol router that selects per-scenario (or per-module) protocols from requirement and runtime signals. ProtocolRouter reduces Fail-Storm recovery time by up to 18.1% versus the best single-protocol baseline, and achieves scenario-specific gains such as higher success in GAIA. We also release ProtocolRouterBench to standardize protocol evaluation and improve reliability at scale. Code and data will be available at: https://github.com/ulab-uiuc/AgentProtocols.

Explore this work

Citation

@misc{protocolbench,
  title = {ProtocolBench: Which LLM MultiAgent Protocol to Choose?},
  author = {Hongyi Du and Jiaqi Su and Jisen Li and Lijie Ding and Yingxuan Yang and Peixuan Han and Xiangru Tang and Kunlun Zhu and Jiaxuan You},
  year = {2026},
  note = {ICML 2026},
  eprint = {2510.17149},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2510.17149}
}
OPEN CONVERSATION /ProtocolBench: Which LLM MultiAgent Protocol to Choose?

Continue the conversation.

Questions and perspectives on this page are welcome.

Prefer a private conversation?

Loading comments…

Leave a public comment

This conversation belongs toProtocolBench: Which LLM MultiAgent Protocol to Choose?. Comments appear only after review. For contact details or personal matters, use the private message form.

Public · reviewed before appearing