Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
New
Hot
Active
Old
Page
1
The Game is Set for a Targeted Memetic Attack on the AI Safety Community
keltan
18 Sep 2026 5:43 UTC
48
points
5
comments
1
min read
LW
link
The Horse
Character#2736
18 Sep 2026 2:52 UTC
25
points
1
comment
3
min read
LW
link
Deep recurrent models are less robustly CoT-monitorable than normal CoT models in a toy setting
Nick Kuhn
and
Alek Westover
18 Sep 2026 2:47 UTC
72
points
0
comments
14
min read
LW
link
Machine intelligence and the death of human expression
Girard Dorney
18 Sep 2026 2:10 UTC
5
points
0
comments
6
min read
LW
link
(extinctiondesk.substack.com)
The Cost of Utopias (a Dialog)
WillPetillo
18 Sep 2026 1:57 UTC
18
points
2
comments
17
min read
LW
link
Two Axes of Alignment: A Framework for Robust Superintelligence Alignment
Arihant Gadgade
18 Sep 2026 0:59 UTC
6
points
0
comments
4
min read
LW
link
Superintelligence this Christmas
Alexander Gietelink Oldenziel
18 Sep 2026 0:06 UTC
46
points
7
comments
3
min read
LW
link
What is (and isn’t) gained by avoiding architectures with high opaque serial depth?
Alek Westover
17 Sep 2026 23:52 UTC
11
points
0
comments
3
min read
LW
link
If METR is overworked, how to alleviate the bottleneck?
Matthew_Opitz
17 Sep 2026 22:56 UTC
16
points
0
comments
3
min read
LW
link
AI is an abundance of choice not a 1D spectrum
KatjaGrace
17 Sep 2026 22:39 UTC
7
points
0
comments
1
min read
LW
link
(worldspiritsockpuppet.substack.com)
AI can kill us without human extinction: P(Catastrophe)
Young Jae Koh
17 Sep 2026 22:22 UTC
3
points
0
comments
4
min read
LW
link
Grantmakers aren’t afraid to die
dan.parshall
17 Sep 2026 21:52 UTC
33
points
14
comments
5
min read
LW
link
(unsolicitedadvice.ai)
Against AI Risk becoming mainstream
Prometheus
17 Sep 2026 21:49 UTC
15
points
3
comments
5
min read
LW
link
The J-lens offset is the model’s token frequency: z-scoring helps
Ameya Panchal
17 Sep 2026 21:13 UTC
7
points
0
comments
10
min read
LW
link
(ameya-bit.github.io)
A Defense of Gradual Disempowerment
Max Harms
17 Sep 2026 21:04 UTC
19
points
3
comments
6
min read
LW
link
Swarm Organization as the Exponent on Test-Time Compute
Julian Bradshaw
17 Sep 2026 20:17 UTC
9
points
1
comment
6
min read
LW
link
Good and bad ways to evaluate a definition
Elijah
17 Sep 2026 18:17 UTC
20
points
0
comments
4
min read
LW
link
Pacing the Frontier: A Framework & Research Agenda
CharlesD
,
technicalities
,
Raymond Douglas
and
Nowe Moore
17 Sep 2026 16:36 UTC
34
points
0
comments
2
min read
LW
link
(pacing.tech)
A Possible Solution to the Observability Problem
Isha Yiras Hashem
17 Sep 2026 14:50 UTC
2
points
0
comments
7
min read
LW
link
Astra uses some of its no-CoT capability in practice
StevenW
17 Sep 2026 14:21 UTC
9
points
2
comments
3
min read
LW
link
Back to top
Next