Score: 1

Goal-Aware Identification and Rectification of Misinformation in Multi-Agent Systems

Published: May 31, 2025 | arXiv ID: 2506.00509v1

By: Zherui Li , Yan Mi , Zhenhong Zhou and more

Potential Business Impact:

Stops fake news from fooling AI teams.

Business Areas:
Semantic Search Internet Services

Large Language Model-based Multi-Agent Systems (MASs) have demonstrated strong advantages in addressing complex real-world tasks. However, due to the introduction of additional attack surfaces, MASs are particularly vulnerable to misinformation injection. To facilitate a deeper understanding of misinformation propagation dynamics within these systems, we introduce MisinfoTask, a novel dataset featuring complex, realistic tasks designed to evaluate MAS robustness against such threats. Building upon this, we propose ARGUS, a two-stage, training-free defense framework leveraging goal-aware reasoning for precise misinformation rectification within information flows. Our experiments demonstrate that in challenging misinformation scenarios, ARGUS exhibits significant efficacy across various injection attacks, achieving an average reduction in misinformation toxicity of approximately 28.17% and improving task success rates under attack by approximately 10.33%. Our code and dataset is available at: https://github.com/zhrli324/ARGUS.

Repos / Data Links

Page Count
21 pages

Category
Computer Science:
Computation and Language