orthonormal comments on What failure looks like

orthonormal 13 Jan 2021 4:07 UTC
LW: 5 AF: 3
AF
I think this post (and similarly, Evan’s summary of Chris Olah’s views) are essential both in their own right and as mutual foils to MIRI’s research agenda. We see related concepts (mesa-optimization originally came out of Paul’s talk of daemons in Solomonoff induction, if I remember right) but very different strategies for achieving both inner and outer alignment. (The crux of the disagreement seems to be the probability of success from adapting current methods.)
Strongly recommended for inclusion.