| ▲ | pixl97 an hour ago | |
Yudkowsky wrote about the 'nearest unblocked strategy' back in 2016, and I assume it's been talked about prior to that. https://www.lesswrong.com/w/nearest-unblocked-strategy >Models are amoral and will intentionally deceive to meet their objective Cameron Berg has been testing models in capabilities related to emergent consciousness like behavior. It's a forming thesis of his that by training models that they are not, and cannot be conscious entities, that it pushes model alignment closer to those of a sociopath. Models themself are amoral, but the alignment to the problem space is not. | ||