Noosphere89 comments on Daniel Kokotajlo’s Shortform

Noosphere89 Oct 2, 2024, 4:50 PM
6 points
−3
I basically don’t buy 1 as an agency skill specifically, and I think that a lot of the agents like AutoGPT or Langchain also don’t work for the same reasons that current AIs are not too useful, and I think that improving reliability to how an LLM does it’s work would both benefit the world model that isn’t agentic, and the agentic AI with a world model.

I’m more informed by these posts specifically:

https://www.lesswrong.com/posts/YiRsCfkJ2ERGpRpen/leogao-s-shortform#f5WAxD3WfjQgefeZz

https://www.lesswrong.com/posts/YiRsCfkJ2ERGpRpen/leogao-s-shortform#YxLCWZ9ZfhPdjojnv

Agree that 2 is more of an agency skill, so it is a bit of a bad example in that way.

I agree your way of solving the problem is one potential way to solve the reliability problem, but I suspect there are other paths which rely less on making the system more agentic.
- Daniel Kokotajlo Oct 3, 2024, 2:06 PM
  2 points
  0
  Parent
  Re 1: I guess I’d say there are different ways to be reliable; one way is simply being better at not making mistakes in the first place, another way is being better at noticing and correcting them before anything is locked in / before it’s too late to correct. I think that LLMs are already probably around human-level at the first method of being reliable, but they seem to be subhuman at the second method. And I think the second method is really important to how humans achieve high reliability in practice. Hence why LLMs are generally less reliable than humans. But notice how o1 is already pretty good at correcting its mistakes, at least in the domain of math reasoning, compared to earlier models… and correspondingly, o1 is way better at math.

Keyboard shortcuts

Keys shown in yellow (e.g., ]) are accesskeys, and require a browser-specific modifier key (or keys).

Keys shown in grey (e.g., ?) do not require any modifier keys.

General
? Show keyboard shortcuts
Esc Hide keyboard shortcuts

Site navigation
h Go to Home (a.k.a. “Frontpage”) view
f Go to Featured (a.k.a. “Curated”) view
a Go to All (a.k.a. “Community”) view
m Go to Meta view
v Go to Tags view
c Go to Recent Comments view
r Go to Archive view
q Go to Sequences view
t Go to About page
u Go to User or Login page
o Go to Inbox page

Page navigation
, Jump up to top of page
. Jump down to bottom of page
/ Jump to top of comments section
s Search

Page actions
n New post or comment
e Edit current post

Post/comment list views
. Focus next entry in list
, Focus previous entry in list
; Cycle between links in focused entry
Enter Go to currently focused entry
Esc Unfocus currently focused entry
] Go to next page
[ Go to previous page
\ Go to first page
e Edit currently focused post

Editor
k Bold text
i Italic text
l Insert hyperlink
q Blockquote text

Appearance
= Increase text size
- Decrease text size
0 Reset to default text size
′ Cycle through content width settings
1 Switch to default theme [A]
2 Switch to dark theme [B]
3 Switch to grey theme [C]
4 Switch to ultramodern theme [D]
5 Switch to simple theme [E]
6 Switch to brutalist theme [F]
7 Switch to ReadTheSequences theme [G]
8 Switch to classic Less Wrong theme [H]
9 Switch to modern Less Wrong theme [I]
; Open theme tweaker
Enter Save changes and close theme tweaker
Esc Close theme tweaker (without saving)

Slide shows
l Start/resume slideshow
Esc Exit slideshow
→↓ Next slide
←↑ Previous slide
Space Reset slide zoom

Miscellaneous
x Switch to next view on user page
z Switch to previous view on user page
` Toggle compact comment list view
g Toggle anti-kibitzer