| ▲ | axus 4 hours ago | |
Article says that there's no human intention or design directing the output we get, but I don't think that's completely right. It's not human, but the algorithm is like the Human Instrumentality Project: an amalgamation of human intentions. That might be more creepy :) | ||
| ▲ | _aavaa_ 2 hours ago | parent [-] | |
> Article says that there's no human intention or design directing the output we get I mean that’s objectively wrong for any model using RLHF. | ||