The authors evaluated five receiver models across five benchmarks.
Incorrect peer messages can override an agent’s evidence-backed answer
Controlled tests across five benchmarks show that inter-agent communication helps with mistakes but can also induce specific wrong answers.
Chinese Tech
Yaxin Gong · Gangyi Zhang · Chongming Gao · Leyang Shen · Chenxiao Fan · Jiakai Wang · +4 more
University of Science and Technology of China · Qwen Business Unit of Alibaba · National University of Singapore
Research Digest··2 min read
Gong and colleagues isolate the effect of message passing in multi-agent LLM systems by holding a downstream agent’s task and evidence fixed while varying the upstream message.
Why this paper
From Qwen Business Unit of Alibaba and 2 others
In one line
In multi-agent LLM systems, incorrect upstream messages override correct downstream answers in up to 32% of cases; 94% of harmful cases copy the upstream's wrong answer.
What we could check
- ·No code link found
- ·No weights link found
- ·No dataset link found
- ·No compute details found
- ·No stated limitations found
- ·No benchmark numbers found
Observed from the paper text and links we have. Absence here means we did not find it, not that it does not exist.
§