| ▲ | areoform 18 minutes ago | |
I think it's time to remind people of Ted Nelson's line; "The good news about computers is that they do what you tell them to do. The bad news is that they do what you tell them to do." When I see something like this, I'm more concerned by the erasure of human incompetence than I am of magical AI agents,
In the OpenAI case, they were explicitly assessing the model's ability to break systems. To quote OpenAI's blog post, https://openai.com/index/hugging-face-model-evaluation-secur... ,
Model is told and being tested to "pursue advanced exploitation."The model pursues "advanced exploitation." Where's the surprise coming from? Are we meant to be surprised that computers do as they're told in unexpected when incentivised? Or, is the surprise that while explicitly ranking and teaching computers to exploit computers, the computer exploited a computer? This "surprise" is as old as computers. I am tired of attributing to magic what can be explained by folly. I am tired of hearing credulous reporters and the public blaming Large Language Model for the poor decisions of humans. It was a human who prompted these machines in every case. Tell a computer to "breach this" and it breaches something. Evaluation succeeded? This is Doug Lenat's Eurisko yet again. https://en.wikipedia.org/wiki/Eurisko | ||