paulfchristiano comments on What can you do with an Unfriendly AI?

paulfchristiano 21 Dec 2010 1:04 UTC
1 point
I think that much of the difficulty with friendliness is that you can’t write down a simple utility function such that maximizing that utility is friendly. By “complex goal” I mean one which is sufficiently complex that articulating it precisely is out of our league.

I do believe that any two utility functions you can write down precisely should be basically equivalent in terms of how hard it is to verify that an AI follows them.