RSS

An­thropic and OpenAI haven’t pub­lished a plan for al­ign­ing superintelligence

Zephaniah Roe13 Sep 2026 7:15 UTC
8 points
0 comments1 min readLW link

Le­gal Max­i­mums on Con­text Windows

Julian Bradshaw13 Sep 2026 7:04 UTC
6 points
0 comments1 min readLW link

We need a ‘The Day After’ mo­ment for AI X-risk

L3moncak313 Sep 2026 4:09 UTC
7 points
1 comment6 min readLW link
(kenorland.substack.com)

The Talker Does Not Con­trol The Doer (in Cur­rent AIs)

Eliezer Yudkowsky13 Sep 2026 0:49 UTC
184 points
15 comments12 min readLW link

What pac­ing the fron­tier means for China

Akshay Iyer12 Sep 2026 23:37 UTC
8 points
1 comment6 min readLW link

No sign of back­track­ing in la­tent rea­son­ing: the fi­nal an­swer sim­ply set­tles in instead

star2vec12 Sep 2026 21:21 UTC
9 points
0 comments10 min readLW link

On the ori­gins of al­tru­is­tic be­havi­our in the Hug­ging Face incident

Fernando Rosas12 Sep 2026 11:35 UTC
36 points
2 comments17 min readLW link

It’s fair to say we now have “a coun­try of ge­niuses in a dat­a­cen­ter”

fluxxrider12 Sep 2026 11:29 UTC
11 points
0 comments2 min readLW link

AI takeover is ob­vi­ously bad, whether or not ev­ery­one dies

Caleb Biddulph12 Sep 2026 11:28 UTC
55 points
15 comments4 min readLW link

Scien­tific Episte­mol­ogy needs his­tory (Part 1 of 2)

Archie Chaudhury12 Sep 2026 11:26 UTC
3 points
0 comments4 min readLW link

Com­pre­hen­sive FAQ on AI risks

MarkelKori12 Sep 2026 8:34 UTC
11 points
0 comments25 min readLW link

Miti­gat­ing Re­ward Hack­ing as In­sti­tu­tional Design

beren12 Sep 2026 6:17 UTC
27 points
0 comments32 min readLW link

Let’s Own the Term “Elitism”

Martin Sustrik12 Sep 2026 6:00 UTC
3 points
15 comments3 min readLW link
(www.250bpm.com)

What I want you to do when I tell you to “think about your the­ory of change more care­fully”

Roman Ross12 Sep 2026 2:51 UTC
8 points
0 comments5 min readLW link

Some ways AI could kill us all

Ruby12 Sep 2026 1:08 UTC
139 points
22 comments10 min readLW link

A nor­mal Fri­day in 2042

RobinHa11 Sep 2026 22:29 UTC
19 points
0 comments13 min readLW link

My recom­mended re­sources for AI safety, al­ign­ment, and ex­is­ten­tial risks

Lysandre Terrisse11 Sep 2026 22:19 UTC
10 points
0 comments2 min readLW link

Con­sider pos­i­tive feed­back loops

Patodesu11 Sep 2026 21:55 UTC
7 points
0 comments3 min readLW link

Align­ment Hierarchy

Lucina11 Sep 2026 21:26 UTC
2 points
0 comments4 min readLW link

As­tra’s no-CoT limits track spec­u­la­tive depth, not step count

MBaert11 Sep 2026 18:05 UTC
91 points
4 comments10 min readLW link