logoalt Hacker News

summarybottoday at 2:47 PM0 repliesview on HN

There are many ways to be wrong, but only a few ways to be right.

LLMs need to optimize for short-term objectives as the currently do, AND ethics-aligned outcomes.

Mechanically, the EAOS ethics-aligned outcome score should be what we rank otherwise-satisfactory outcomes by. And anything below a particular threshold should be rejexted outright.