LessWrong
techCenter
Variance of Valuetranslating…
Here is a question worth asking at least once: Why can't we just solve alignment by doing RL where the reward is exactly equal to our own utility function?
Now, there are some implementation concerns here. For example, we don't actually know our own utility function. And even if…
Keywords#Value#Variance#own utility#utility function#tech