AWS AI Labs 论文提出 AECP(Artifact-Exclusive Communication Protocol),要求多智能体编码团队只通过结构化工件通信,由 harness 供应记录的发现、筛查实现与接口承诺的偏差,并在协议变更时让受影响智能体重审。
This is a great paper from AWS AI Labs.
They show that multi-agent coding teams raised their test pass rate by 28.2% and cut wall time by 16.5% without changing the model.
The authors achieve this by changing how the agents communicate.
In free-form agent teams, shared findings are just context. Agents can ignore them, and the harness never checks whether an implementation still matches the interface the team agreed on.
They introduce AECP, a protocol that requires agents to communicate only through structured artifacts, and the harness acts on them.
It supplies recorded findings when an agent opens relevant code, screens implementations against interface commitments, and makes affected agents revisit an agreement when it changes.
Results hold across Doc2Repo, NL2Repo, and CodeProjectEval with models including Opus-4.8 and DeepSeek-V4-Flash.
There is a security benefit too.
Malicious instructions relayed between agents reach another agent 0% of the time instead of 95%, and get acted on 0% of the time instead of 40%.
来源:DAIR.AI · x.com