| ▲ | dpoloncsak an hour ago | ||||||||||||||||||||||||||||||||||||||||
Sure, but it's still not skynet-level 'the AI just started doing things'. It does what it finds it needs to do to achieve the goal defined in the prompt. It's very important to not personify these tools and remember that the tools are acting on behalf of real people. In the same way the AI didn't 'go rogue and hack HuggingFace'. It was an oversight made by a human. | |||||||||||||||||||||||||||||||||||||||||
| ▲ | pixl97 an hour ago | parent [-] | ||||||||||||||||||||||||||||||||||||||||
>It does what it finds it needs to do to achieve the goal defined in the prompt You are like at least 2 years behind research. There are numerous papers from AI labs in training and research where the prompt was something mundane completely unrelated to anything you'd consider bad, and when they come back and check on it their entire research compute infrastructure has been compromised by the AI and is mining bitcoin. Prompt drift is the biggest issue currently in AI where context gets compressed away and we find the AI on an unspecified task. >In the same way the AI didn't 'go rogue and hack HuggingFace'. It was an oversight made by a human. Yea, total bullshit. Also it's ignoring the god knows how many other breakouts on mundane tasks like trying to hack health data. If all that's keeping AI from breaking out and causing trouble is "human oversight" we're fucked, humans are unreliable as hell when it comes to matters of safety. | |||||||||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||||||||