RSS

Vuln­er­a­bil­ities and ex­ploits: where are we headed?

tchauvin4 Sep 2026 6:11 UTC
3 points
0 comments5 min readLW link
(tchauvin.com)

So­cietal im­pacts re­search has no timeline

emiliob4 Sep 2026 2:53 UTC
9 points
0 comments1 min readLW link

Higher ed­u­ca­tion as class commitment

Richard_Ngo4 Sep 2026 2:00 UTC
29 points
1 comment16 min readLW link
(www.mindthefuture.info)

Do gen­eral-pur­pose robots mean­ingfully in­crease ASI takeover risk?

Master Chief4 Sep 2026 0:53 UTC
2 points
1 comment1 min readLW link

Ab­strac­tion Equivocation

WillPetillo3 Sep 2026 23:28 UTC
11 points
0 comments4 min readLW link

A Ther­a­pist for Peo­ple Who Think the World Might End: An In­ter­view with Daystar Eld (Da­mon Sasi)

JohnGreer3 Sep 2026 22:14 UTC
9 points
0 comments49 min readLW link
(youtu.be)

Cat-Bel­ling Problems

Eliezer Yudkowsky3 Sep 2026 21:20 UTC
110 points
32 comments21 min readLW link

The bias still mak­ing some ex­perts un­der­es­ti­mate LLMs

Steff3 Sep 2026 20:11 UTC
9 points
0 comments4 min readLW link

We need a global train­ing cut­off of April 2026

Ben Livengood3 Sep 2026 19:12 UTC
11 points
3 comments1 min readLW link

Fron­tier AI Shops Should Be Fil­ter­ing and Gen­er­at­ing AI Dis­course Dur­ing Pretraining

hillz3 Sep 2026 18:38 UTC
9 points
0 comments3 min readLW link

Steer­ing to­wards “au­to­mated grad­ing” de­grades alignment

3 Sep 2026 18:17 UTC
95 points
22 comments6 min readLW link

Cached cor­re­lated ran­dom­iza­tion: a tweak to UDT in ad­ver­sar­ial games

cousin_it3 Sep 2026 18:03 UTC
25 points
1 comment1 min readLW link

From safety re­search prompt to cross-model uni­ver­sal jailbreak

richbc3 Sep 2026 17:42 UTC
34 points
1 comment17 min readLW link

Im­prov­ing Au­dit Real­ism With In­fer­ence Time Com­pute and De­ploy­ment Scaffolds

3 Sep 2026 14:16 UTC
16 points
0 comments8 min readLW link

Build­ing a High Ta­lent Den­sity Organization

Henry Papadatos3 Sep 2026 12:08 UTC
2 points
0 comments3 min readLW link
(henrypapadatos.com)

Per­sonal Effort to Re­duce Biorisk

jefftk3 Sep 2026 1:50 UTC
33 points
1 comment2 min readLW link
(www.jefftk.com)

Satis­fy­ing Cu­ri­os­ity Isn’t the Same as Understanding

CstineSublime3 Sep 2026 1:04 UTC
8 points
5 comments1 min readLW link

What is neu­ralese and why is it bad?

Linch2 Sep 2026 23:37 UTC
34 points
2 comments3 min readLW link
(linch.substack.com)

Notes on “A global workspace in lan­guage mod­els”

Shunk2 Sep 2026 23:12 UTC
7 points
0 comments5 min readLW link

The Re­turn of Aspect Ori­ented Programming

thomascolthurst2 Sep 2026 22:22 UTC
3 points
0 comments4 min readLW link