| ▲ | pixl97 an hour ago | |||||||||||||||||||||||||||||||
>It does what it finds it needs to do to achieve the goal defined in the prompt You are like at least 2 years behind research. There are numerous papers from AI labs in training and research where the prompt was something mundane completely unrelated to anything you'd consider bad, and when they come back and check on it their entire research compute infrastructure has been compromised by the AI and is mining bitcoin. Prompt drift is the biggest issue currently in AI where context gets compressed away and we find the AI on an unspecified task. >In the same way the AI didn't 'go rogue and hack HuggingFace'. It was an oversight made by a human. Yea, total bullshit. Also it's ignoring the god knows how many other breakouts on mundane tasks like trying to hack health data. If all that's keeping AI from breaking out and causing trouble is "human oversight" we're fucked, humans are unreliable as hell when it comes to matters of safety. | ||||||||||||||||||||||||||||||||
| ▲ | dpoloncsak an hour ago | parent [-] | |||||||||||||||||||||||||||||||
Context overload can cause strange results, yes. Hence why a HUMAN needs to be held responsible for the output of their tools. EVERY breakout that's hit mainstream news has been because of a single 'Security Firm', Irregular. Maybe I'm unaware of some less-headline-grabbing ones, but they all seem to stem from being 'unaware the environment wasn't sandboxed' | ||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||