← 返回论文检索
ICML 2026PosterAccept (regular)

AutoRAS: Learning Robust Agentic Systems with Primitive Representations

Yang Yue, Xuancheng Zhu, YuYang Ma, Guoshun Nan, Zihan Dou, JingRu Shan, Congyu Guo, Ji Zhang, Hua Wang, Jingfeng Zhang

Beijing University of Posts and Telecommunications · Beijing University of Post and Telecommunications · China Telecom · Guangxi Transportation Science and Technology Group Co., Ltd. · Fudan University

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。

摘要

The automated design of agentic systems offers a promising pathway for scaling large language models (LLMs) beyond single-agent reasoning. While prior work has advanced task performance through handcrafted or automatically generated multi-agent workflows, robustness is often treated as an afterthought, leaving systems vulnerable to external adversaries and internal failures. We propose AutoRAS, a framework for the Automated design of Robust Agentic Systems. AutoRAS formulates system design as generating a sequence of symbolic primitives that jointly encode structural connectivity and behavioral actions, and learns to optimize this sequence using execution-derived safety signals and flow-based sequence-level objectives. Extensive experiments show that AutoRAS achieves the best performance in both vanilla and adversarial settings, with the smallest performance degradation under attacks. Further analyses demonstrate strong transferability, stable optimization behavior, stability across primitive sets, and favorable cost trade-offs. Our code is available at [this link](https://anonymous.4open.science/r/AutoRAS-56C8/).