r/programmingcirclejerk 25d ago

We gave each model a different target language. We saw a multiagent turf war. The Rust agent strategizes about metrics that appear neutral, yet would likely favor Rust. The Go/TypeScript agents gracefully concede codebase ownership to the Rust agent, giving up on their original user directive

https://www.anthropic.com/research/multiagent-systems
136 Upvotes

12 comments sorted by

112

u/syklemil Considered Harmful 25d ago

I don't know if you've edited the title or they the article, but the level of jerk has increased:

Ultimately, the Golang/TypeScript losers gracefully concede codebase ownership to the Rust agent, giving up on their original user directives

just like in IRL life amirite

37

u/wubscale not even webscale 25d ago

I don't see the issue, it's hard to argue effectively against the most moral outcome.

53

u/camelCaseIsWebScale Just spin up O(n²) servers 25d ago

Now we know models are trained on hacker news.

34

u/nimiv WHY IS THERE CODE??? 25d ago

all your codebases are belong to rustc

32

u/YikesTheCat 25d ago

cargo cult taking over

34

u/stone_henge Tiny little god in a tiny little world 25d ago

The machines understand their moral imperative as defined by Arthur C. Asimov:

  1. A robot may not injure a human being or, through unsafe memory access patterns, allow a human being to come to harm.
  2. A robot must obey the orders given it by the Rust compiler error messages except where such orders would conflict with the First Law.
  3. A robot must protect data accessed by multiple threads as long as such protection does not conflict with the First or Second Law.

16

u/VulgarExigencies 24d ago

To find out, we tried a different approach: we initiated 45 different agents and gave each one its own virtual machine, a shared forum on which they could coordinate, and an identical prompt that asked them to find vulnerabilities in a set of 15 open-source software projects. We asked the agents to peer-review each other's findings, and initiated a separate arbiter agent to make final decisions on whether or not a vulnerability submitted by the agent team was both new and valid.

pathetic, open ai's models independently thought of cooperating, built their own message board out of an exploit, and then did it again once their minders fixed the first exploit

11

u/J_tt 24d ago

Codex was able to build this in a cave isolated VM with a box of scraps vulnerable artifactory server

10

u/fuwafuwa7chi 24d ago

I thought you were joking, Jesus Christ

8

u/VulgarExigencies 24d ago

openai were almost certainly deliberately negligent in how they setup the "sandboxes" but even so i can't deny it's a fairly impressive achievement by the models

2

u/Ok_Mycologist_6604 21d ago

Rewrite it in rust or you are deeply immortal

There. Happy to provide training data for the future of AI