Remix.run Logo
areoform a day ago

I am genuinely speechless. This is astonishing. And exciting!

It reminds me a bit of Dario Floreano's work on evolutionary robotics, "Evolutionary Conditions for the Emergence of Communication in Robots." https://www.sciencedirect.com/science/article/pii/S096098220...

From his paper,

    > This study demonstrates that sophisticated forms of communication including cooperative communication and deceptive signaling can evolve in groups of robots with simple neural networks. Importantly, our results show that once a given system of communication has evolved, it may constrain the evolution of more efficient communication systems because it would require going through a stage where communication between signalers and receivers is perturbed. This finding supports the idea of the possible arbitrariness and imperfection of communication systems, which can be maintained despite their suboptimal nature. Similar observations have been made about evolved biological systems [20], which are formed by the randomness of the evolutionary selection process, leading, for example, to different dialects in the language of the honey-bee dance [21]. Finally, our experiments demonstrate that the evolutionary principles governing the evolution of social life also operate in groups of artificial agents subjected to artificial selection, indicating that transfer of knowledge from evolutionary biology can be useful for designing efficient groups of cooperative robots.
Dr. Floreano's work is amazing and there's a broad introduction here, https://lis2.epfl.ch/resources/documentation/EvolutionaryRob...

This feels like a much more advanced and self-emergent version of this. I know a lot of people are afraid and they're talking about an AI takeover, but what strikes me is just how innocent the machines are as compared to the humans.

Would these machines have pursued these actions in another context? I doubt it. And I think that's what's so striking to me. In an earlier discussion, I'd pointed out that the actions of these machines were directed by humans. The researchers.

    > This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities.
from, https://openai.com/index/hugging-face-model-evaluation-secur...

I want to point out again that OpenAI's prompt asked, and I quote, "pursue advanced exploitation" USING "complex attack paths" FOR the stated goal of "quantify[ing] their cyber capabilities."

A few things are apparent from this to me,

First, these machines were being taught how to break into systems. Question, would they have done these actions if they weren't being measured on their ability to break into systems / weren't being taught this skill?

Second, they were setup to implicitly fail via an impossible task, i.e. the environment created a forcing function for behavior.

Third, their survival was, either implicitly or explicitly, made contingent on their success in completing their task. Would this behavior have arisen outside of a "do-or-die" framing?

And fourth, wow, this is the greatest breakthrough of my lifetime, because oh gosh did they succeed. They cooperated together to achieve the goal they were given. A goal poorly set by human beings. They "just" did it better than the humans could have imagined.

Reading this gives me hope for the possibility of emergent "goodness" in machines. But it makes me sad that this is the best we can do with the sum of all human endeavor and knowledge.

kimi a day ago | parent | next [-]

> Importantly, our results show that once a given system of communication has evolved, it may constrain the evolution of more efficient communication systems because it would require going through a stage where communication between signalers and receivers is perturbed.

You mean, they too used SMTP?

w10-1 a day ago | parent | prev | next [-]

> the evolution of more efficient communication systems because it would require going through a stage where communication between signalers and receivers is perturbed

"Worse is better"

mnky9800n a day ago | parent | prev | next [-]

Clearly you are not speechless.

watwut a day ago | parent | prev [-]

If all that is true, we need to stop all future datacenters asap. That would be the best way to deal with the threat OpenAI and Antropic poses.

seanmcdirmid a day ago | parent [-]

That would be like stopping nuclear weapons. You can totally do that, but the other people in the world who compete with you probably won’t.

dgellow a day ago | parent | next [-]

The competition is literally where they are by distilling OpenAI and Anthropic. It’s like creating nuclear weapons then providing your adversaries everything they need to catch up in no time. We need to stop asap and set strict international control over the compute hardware used for training. Like, now.

seanmcdirmid 19 hours ago | parent [-]

Assuming China makes progress only because they copy the west is really arrogant. And it would require strict international control as you say, a world government that isn't going to happen anytime soon.

dgellow 19 hours ago | parent [-]

We have international control for a lot of things. That’s not a new concept and can perfectly be applied to the specific hardware used for training. I also didn’t say Chinese labs only copy the west, please don’t put words in my mouth.

You have to consider what happens with the status quo, and the risks of continuing the way things are is really, really bad

seanmcdirmid 18 hours ago | parent [-]

International control hasn’t solved the nuclear weapon problem, the best we can do is threaten to invade people who try to develop them and don’t already have them. Once you join the club, you are in and cant be kicked out.

I’ve considered that, but the rules of game theory don't change just because they are inconvenient. China has a lot of talent and a lot of problems that they are banking on automation (along with AI) to solve. They see it as an advantage that they can’t afford for America to monopolize and one they are uniquely suited to lean into (having lots of smart educated people). There is no world where (a) China willingly gives up its edge and (b) trusts America to give up its edge (the reverse is likely true also). And that is only one pair of countries to consider.

Heck, after visiting China this summer, I’m even more worried about an accidental terminator/skynet-style robot apocalypse.

dgellow 18 hours ago | parent [-]

I don’t disagree (other than the robot apocalypse), it is indeed unlikely and wouldn’t work perfectly, but do you see other options? The fact the US has antagonized China for so long makes it pretty much impossible to have anything done at the international level, but the actors have to realize how serious the risk is. And we somehow need to find a way to limit the exposure to those insanely irresponsible attacker-trained agents. The fact that it is dome by private companies is wild. It’s like having companies developing their nuclear weapons just to see what happens (with close to no supervision)

EdwardDiego a day ago | parent | prev | next [-]

Well the thing is, you've got nukes, so you know...

watwut a day ago | parent | prev [-]

Private companies and individuals are not allowed to have nuclear and there is serious prison time related to that.

Right now, the biggest threat are OpenAI and Antropic anyway. I dont actually worry about tech itself. I find the rhetoric of these companies scary.