For the purposes of this argument, I’m interested in what can be done safely by some AI we can build. If you can solve alignment safely with some AI, then you’re in a good situation. What an arbitrarily powerful optimiser will do isn’t the crux, we all agree that’s dangerous.
For the purposes of this argument, I’m interested in what can be done safely by some AI we can build. If you can solve alignment safely with some AI, then you’re in a good situation. What an arbitrarily powerful optimiser will do isn’t the crux, we all agree that’s dangerous.
I see what you’re getting at. Interesting question.