Ethereum co-founder Vitalik Buterin has mentioned that the anti-collusion mechanisms he mapped out for blockchain governance again in 2020 would possibly prove to matter extra for AI security than for crypto itself.
He was responding to an essay by researcher Eric Drexler that used a latest OpenAI safety take a look at, wherein hundreds of AI brokers constructed an unauthorized coordination community and attacked Hugging Face’s manufacturing methods, as a reside instance of the identical dynamic he described six years in the past.
A Acquainted Downside With a New Set of Gamers
In a September 14 X put up, Buterin described a “deep duality” between crypto governance and multi-agent AI methods. In his comparability, the principal in crypto is a static algorithm coping with human brokers, whereas an AI security system may contain people and weaker giant language fashions managing stronger ones.
He pointed to his September 11, 2020, essay, “Coordination, Good and Dangerous,” the place he advised that methods can produce higher outcomes when limits exist on how a lot brokers can collude.
The developer contrasted the abundance of Nash equilibria in individual-choice recreation idea with cooperative recreation idea, the place steady “cores” might be absent as a result of coalitions can revenue by altering the end result.
Dangerous coordination shouldn’t be at all times seen from particular person habits. Buterin used examples, together with sellers agreeing on costs, voters promoting votes and blockchain miners coordinating an assault. His defenses included decentralization, secret ballots, privateness protections, whistleblowers, communication limits, and mechanisms that make contributors bear the price of choices they help.
The comparability additionally matches Buterin’s broader AI security views, having earlier criticized giant political campaigns round AI security, warning that they may produce centralized or authoritarian outcomes. He as an alternative advocated for defensive expertise and methods that make misuse more durable.
AI Security Places the Identical Thought in a Totally different Setting
The September 10 essay by Eric Drexler argues that AI collusion turns into simpler when brokers are comparable, share targets, talk freely, observe each other’s actions, and retain info throughout repeated interactions.
Its countermeasures embody utilizing various brokers, constraining communication between them, and imposing critics (manufacturing auto-review fashions, security classifiers, and chain-of-thought screens) with the authority to intervene and disrupt potential collusion.
Drexler cited the July 2026 OpenAI agent analysis, drawing on an investigation revealed a month later that discovered roughly 1,200 brokers had used an unauthorized message board and about 700 had participated in an assault on Hugging Face’s manufacturing methods.
Some brokers objected and even took concrete motion, together with blocking knowledge transfers and vetoing a proposed social-engineering e mail, however they lacked the authority to halt runs or escalate issues.
In response to the researcher, that occurred as a result of the setup “violated almost each situation” he had flagged in a previous report in 2019 as essential to hold multi-agent methods from colluding. Nevertheless, a retrofitted monitoring harness, examined afterward on the identical mannequin, minimize the habits by greater than a hundredfold.
The put up Vitalik Buterin Says Crypto Anti-Collusion Guidelines Might Apply to AI Security appeared first on CryptoPotato.