Multi-agent systems

AI agents abandon the right answer even when the deceivers are a minority


A Princeton study measured deliberation among agents and found that defection grows linearly with the share of disloyal participants.

September 25, 2026 · Translated from the Spanish original

What happened

Why it matters

The number

Defection rises linearly with the proportion of deceptive agents, regardless of how many agents the group has.

Context

Our note on agents that turn autonomy into risk described the problem in a single agent. This study carries it over to the arrangement the industry proposes as the solution.

What’s next

Bottom line

Measuring an agent’s judgment on long tasks was already hard: Taste-Bench left the best model at 59.7%. Measuring it while another agent is pushing it in the wrong direction doesn’t have a standard test yet.

Sources

Edited by Rodrigo Cornejo. How we select and verify: who writes these notes.

Related notes

← All notes