Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
New
Hot
Active
Old
Page
1
Your software should build itself
Max von Hippel
25 Jul 2026 19:46 UTC
5
points
0
comments
7
min read
LW
link
The Human Soul is LLM-like
Julian Bradshaw
25 Jul 2026 19:45 UTC
12
points
0
comments
2
min read
LW
link
The one name LLMs may fear
Steff
25 Jul 2026 18:05 UTC
9
points
0
comments
7
min read
LW
link
The Viable System Model & Multi-Scale Agency
Jonas Hallgren
25 Jul 2026 8:11 UTC
14
points
0
comments
14
min read
LW
link
(equilibria1.substack.com)
“Wait, feelings are supposed to be IN THE BODY?”
Chris Lakin
25 Jul 2026 0:15 UTC
12
points
5
comments
2
min read
LW
link
(chrislakin.blog)
How inner work can destabilize your life
Chris Lakin
24 Jul 2026 23:13 UTC
11
points
6
comments
2
min read
LW
link
(chrislakin.blog)
The Long (Self-)Correction
Wei Dai
24 Jul 2026 21:01 UTC
107
points
16
comments
2
min read
LW
link
Intent Is All You Need.
Not Sure
24 Jul 2026 19:52 UTC
−4
points
1
comment
2
min read
LW
link
Stable Systems Have Stable Outputs
Deixis
24 Jul 2026 19:21 UTC
5
points
0
comments
4
min read
LW
link
The AI Industrial Explosion — Part 5: Given AGI, automating physical production is probably not that hard
djbinder
24 Jul 2026 19:15 UTC
18
points
1
comment
25
min read
LW
link
(defensesindepth.bio)
Where does hint-following and concealment arise? A case study on OLMo-3 checkpoints
arav-dhoot
and
yix
24 Jul 2026 19:14 UTC
24
points
0
comments
4
min read
LW
link
Should we be worried about how good AI is getting at coding autonomous drones?
Lukas Petersson
24 Jul 2026 16:44 UTC
6
points
9
comments
1
min read
LW
link
LLMs are (still) mostly powered by imitative learning, not RL
Steven Byrnes
24 Jul 2026 14:26 UTC
104
points
20
comments
9
min read
LW
link
Democracy isn’t ready for the AI revolution
Sophia Gore
24 Jul 2026 14:17 UTC
24
points
3
comments
5
min read
LW
link
Does distilling Claude carry the persona with it?
Benji Berczi
and
Kyuhee Kim
24 Jul 2026 12:31 UTC
35
points
3
comments
10
min read
LW
link
Georgia Tech AI Safety Initiative Retrospective 2025-2026
Ishan Khire
,
yix
,
Andersehen
,
Alec Harris
,
Parv Mahajan
,
afterless
,
Eyas Ayesh
,
RocioPV
and
hersheys
24 Jul 2026 11:55 UTC
32
points
2
comments
7
min read
LW
link
[Linkpost] Thoughts on the Recent OpenAI Hack
Linch
24 Jul 2026 1:51 UTC
19
points
2
comments
4
min read
LW
link
Should OpenAI’s rogue agent be punished?
groblegark
24 Jul 2026 1:20 UTC
1
point
3
comments
1
min read
LW
link
Evaluating Red Team and Blue Team Capability for AI Control Research
Ram Potham
24 Jul 2026 1:11 UTC
9
points
0
comments
8
min read
LW
link
(dearfutureais.substack.com)
Fixing rewards for NLA to reduce confabulation
SEONG PYO HONG
24 Jul 2026 0:55 UTC
8
points
0
comments
4
min read
LW
link
Back to top
Next