This is a very exciting project! I’m particularly glad to see two features: (i) the focus on “deception”, which undergirds much existential risk but has arguably been less of a focal point than “agency”, “optimization”, “inner misalignment”, and other related concepts, (ii) the ability to widen the bottleneck of upskilling novice AI safety researchers who have, say, 500 hours of experience through the AI Safety Fundamentals course but need mentorship and support to make their own meaningful research contributions.
This is a very exciting project! I’m particularly glad to see two features: (i) the focus on “deception”, which undergirds much existential risk but has arguably been less of a focal point than “agency”, “optimization”, “inner misalignment”, and other related concepts, (ii) the ability to widen the bottleneck of upskilling novice AI safety researchers who have, say, 500 hours of experience through the AI Safety Fundamentals course but need mentorship and support to make their own meaningful research contributions.