Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
New
Hot
Active
Old
Page
1
Automting AI Safety Research: Managing expectations
Gerard Boxo
14 Sep 2026 17:56 UTC
4
points
0
comments
4
min read
LW
link
OpenAI President Brockman says HuggingFace incident model had not been alignment-trained
Caspar Oesterheld
14 Sep 2026 15:28 UTC
26
points
8
comments
1
min read
LW
link
(www.bloomberg.com)
Threat Models for Catastrophic Risks from Decentralised Agent Swarms
Stephen Elliott
14 Sep 2026 15:22 UTC
9
points
0
comments
33
min read
LW
link
[Cross-post] Palisade Podcast episode “How to Actually Influence AI Policy (No Law Degree Required) — with Matthew Lipka”
davekasten
14 Sep 2026 13:40 UTC
27
points
0
comments
55
min read
LW
link
(palisaderesearch.org)
What we have is not what we prepared for
PeacockOfJuno
14 Sep 2026 13:03 UTC
4
points
1
comment
5
min read
LW
link
Current alignment techniques might be ineffective (and actively bad) in the age of RL
Daniel Tan
14 Sep 2026 9:28 UTC
92
points
1
comment
6
min read
LW
link
There is a channel to 900M weekly users. What goes in it?
Charbel-Raphaël
14 Sep 2026 9:21 UTC
119
points
6
comments
3
min read
LW
link
Watch AI materials-science & bioscience abilities closely
Yair Halberstadt
14 Sep 2026 7:10 UTC
14
points
7
comments
1
min read
LW
link
Yet another message board
ErnestScribbler
14 Sep 2026 7:08 UTC
−1
points
0
comments
1
min read
LW
link
RSI will not continue indefinitely
dominicq
14 Sep 2026 6:43 UTC
9
points
8
comments
1
min read
LW
link
(blog.d11r.eu)
Talking points for AI doomers
hgnathan
14 Sep 2026 0:46 UTC
1
point
1
comment
6
min read
LW
link
Deployment
Nina Panickssery
14 Sep 2026 0:00 UTC
66
points
1
comment
3
min read
LW
link
Alignment & Succession: The Two Bars of Alignment
L Rudolf L
13 Sep 2026 20:04 UTC
61
points
7
comments
14
min read
LW
link
Consider how your global governance proposal is different from the EU Code of Practice
David Matolcsi
13 Sep 2026 16:15 UTC
72
points
10
comments
4
min read
LW
link
Teleoperated Humans
jefftk
13 Sep 2026 14:00 UTC
71
points
11
comments
6
min read
LW
link
(www.jefftk.com)
A helpful alignment gadget
Logan Zoellner
13 Sep 2026 13:38 UTC
14
points
3
comments
3
min read
LW
link
Anthropic and OpenAI haven’t published a plan for aligning superintelligence
Zephaniah Roe
13 Sep 2026 7:15 UTC
30
points
12
comments
1
min read
LW
link
Legal Maximums on Context Windows
Julian Bradshaw
13 Sep 2026 7:04 UTC
3
points
5
comments
1
min read
LW
link
We need a ‘The Day After’ moment for AI X-risk
L3moncak3
13 Sep 2026 4:09 UTC
16
points
3
comments
6
min read
LW
link
(kenorland.substack.com)
The Talker Does Not Control The Doer (in Current AIs)
Eliezer Yudkowsky
13 Sep 2026 0:49 UTC
377
points
50
comments
12
min read
LW
link
Back to top
Next