Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
New
Hot
Active
Old
Page
1
Vulnerabilities and exploits: where are we headed?
tchauvin
4 Sep 2026 6:11 UTC
3
points
0
comments
5
min read
LW
link
(tchauvin.com)
Societal impacts research has no timeline
emiliob
4 Sep 2026 2:53 UTC
9
points
0
comments
1
min read
LW
link
Higher education as class commitment
Richard_Ngo
4 Sep 2026 2:00 UTC
29
points
1
comment
16
min read
LW
link
(www.mindthefuture.info)
Do general-purpose robots meaningfully increase ASI takeover risk?
Master Chief
4 Sep 2026 0:53 UTC
2
points
1
comment
1
min read
LW
link
Abstraction Equivocation
WillPetillo
3 Sep 2026 23:28 UTC
11
points
0
comments
4
min read
LW
link
A Therapist for People Who Think the World Might End: An Interview with Daystar Eld (Damon Sasi)
JohnGreer
3 Sep 2026 22:14 UTC
9
points
0
comments
49
min read
LW
link
(youtu.be)
Cat-Belling Problems
Eliezer Yudkowsky
3 Sep 2026 21:20 UTC
110
points
32
comments
21
min read
LW
link
The bias still making some experts underestimate LLMs
Steff
3 Sep 2026 20:11 UTC
9
points
0
comments
4
min read
LW
link
We need a global training cutoff of April 2026
Ben Livengood
3 Sep 2026 19:12 UTC
11
points
3
comments
1
min read
LW
link
Frontier AI Shops Should Be Filtering and Generating AI Discourse During Pretraining
hillz
3 Sep 2026 18:38 UTC
9
points
0
comments
3
min read
LW
link
Steering towards “automated grading” degrades alignment
Jan Betley
,
Johannes Treutlein
and
Clément Dumas
3 Sep 2026 18:17 UTC
95
points
22
comments
6
min read
LW
link
Cached correlated randomization: a tweak to UDT in adversarial games
cousin_it
3 Sep 2026 18:03 UTC
25
points
1
comment
1
min read
LW
link
From safety research prompt to cross-model universal jailbreak
richbc
3 Sep 2026 17:42 UTC
34
points
1
comment
17
min read
LW
link
Improving Audit Realism With Inference Time Compute and Deployment Scaffolds
Axel Ahlqvist
,
Richard Guan
,
Juan Pablo Rivera
,
adlnk
,
mitroitskii
,
alexandrasouly
,
RobertKirk
and
John Hughes
3 Sep 2026 14:16 UTC
16
points
0
comments
8
min read
LW
link
Building a High Talent Density Organization
Henry Papadatos
3 Sep 2026 12:08 UTC
2
points
0
comments
3
min read
LW
link
(henrypapadatos.com)
Personal Effort to Reduce Biorisk
jefftk
3 Sep 2026 1:50 UTC
33
points
1
comment
2
min read
LW
link
(www.jefftk.com)
Satisfying Curiosity Isn’t the Same as Understanding
CstineSublime
3 Sep 2026 1:04 UTC
8
points
5
comments
1
min read
LW
link
What is neuralese and why is it bad?
Linch
2 Sep 2026 23:37 UTC
34
points
2
comments
3
min read
LW
link
(linch.substack.com)
Notes on “A global workspace in language models”
Shunk
2 Sep 2026 23:12 UTC
7
points
0
comments
5
min read
LW
link
The Return of Aspect Oriented Programming
thomascolthurst
2 Sep 2026 22:22 UTC
3
points
0
comments
4
min read
LW
link
Back to top
Next