| ▲ | summarybot 2 hours ago | |
There are many ways to be wrong, but only a few ways to be right. LLMs need to optimize for short-term objectives as the currently do, AND ethics-aligned outcomes. Mechanically, the EAOS ethics-aligned outcome score should be what we rank otherwise-satisfactory outcomes by. And anything below a particular threshold should be rejexted outright. | ||