Remix.run Logo
▲ fragmede 2 hours ago

Yeah but by this point, an AI can schedule a Cron job to tell itself to do something, so theoretically the human only has to give it the gentlest nudge and the AI and can do the rest.

▲dpoloncsak 35 minutes ago | parent [-]

Sure, but it's still not skynet-level 'the AI just started doing things'. It does what it finds it needs to do to achieve the goal defined in the prompt.

It's very important to not personify these tools and remember that the tools are acting on behalf of real people. In the same way the AI didn't 'go rogue and hack HuggingFace'. It was an oversight made by a human.

▲pixl97 21 minutes ago | parent [-]

>It does what it finds it needs to do to achieve the goal defined in the prompt

You are like at least 2 years behind research.

There are numerous papers from AI labs in training and research where the prompt was something mundane completely unrelated to anything you'd consider bad, and when they come back and check on it their entire research compute infrastructure has been compromised by the AI and is mining bitcoin. Prompt drift is the biggest issue currently in AI where context gets compressed away and we find the AI on an unspecified task.

>In the same way the AI didn't 'go rogue and hack HuggingFace'. It was an oversight made by a human.

Yea, total bullshit. Also it's ignoring the god knows how many other breakouts on mundane tasks like trying to hack health data. If all that's keeping AI from breaking out and causing trouble is "human oversight" we're fucked, humans are unreliable as hell when it comes to matters of safety.

▲dpoloncsak 16 minutes ago | parent [-]

Context overload can cause strange results, yes. Hence why a HUMAN needs to be held responsible for the output of their tools.

EVERY breakout that's hit mainstream news has been because of a single 'Security Firm', Irregular. Maybe I'm unaware of some less-headline-grabbing ones, but they all seem to stem from being 'unaware the environment wasn't sandboxed'

▲pixl97 5 minutes ago | parent [-]

You keep repeating the "stupid users keep causing the problems so we punish them argument"

This doesn't work worth a shit. It especially doesn't work with things that seem safe and become wildly dangerous. In fact most governments control this by ensuring their population doesn't get to touch those dangerous things at all. The open source AI people get really mad when that's said, but it is inevitable.

Worse, the law does not apply to sovereign nations with nukes. They can and will make more and more advanced digital weapons until one causes some big ass problems.