| ▲ | kennywinker 3 hours ago | |
My position is that cheating is too slippery a concept to train out. But hey, I am no expert, so maybe I am wrong there. But I'm pretty confidant morality is too slippery a concept to train in. As someone else in these comments said: it's context dependent. As an example: it's wrong to hack the government, right? It's illegal for sure. So we should train AI to follow all the laws. Now what if the government is committing a genocide? Now is it wrong to hack the government? If we just do the first, we get a good nazi soldier. If we train the second as well, maybe we get an oscar schindler. But now we have a model that can be fooled into doing a hack, if it believes that it's for the greater good. So we train it to not be gullible, but now it can't be convinced to help hack even when it's an ethical hack. Too complex, too slippery. Humans fail this stuff all the time. | ||