Remix.run Logo
andsoitis 3 hours ago

Opening line:

> A core hope for managing AI risks is that AIs will help us understand our situation

Gonna stop you right there and ask that you think deeply about that premise.

xlayn 35 minutes ago | parent | next [-]

This should be the top comment, the implication is: AI will decide if AI is correct or no, so adding a new layer of approval...

Whenever you say, hey that's incorrect, the legal process goes to AI and will decide on that issue...

Will work as amazing as the chatbots of the AI companies to solve your issues...

blovescoffee 2 hours ago | parent | prev | next [-]

The underlying idea is that AI capabilities will become so advanced that only AI will enable us to monitor/correct/understand behavior. Obviously this is not without issue and I don't want to try to defend their position right now. But that's what they mean

hn_throwaway_99 2 hours ago | parent | next [-]

That first sentence of yours explains exactly why it is so ridiculous. If only AI can understand it, how is there any assurance that AI will "correct" it's behavior that is aligned with what humans presumably want.

VeninVidiaVicii 2 hours ago | parent | next [-]

But also what is the evidence that something only AI can understand even “matters” or makes sense? I’m increasingly convinced commercial AI is exploiting our logical blind spot to be spoken to authoritatively.

skybrian 2 hours ago | parent | prev | next [-]

It’s not hard to get LLM’s to inform on each other. They don’t really do loyalty.

vhantz 31 minutes ago | parent [-]

It's a language model. It generates text.

esafak 2 hours ago | parent | prev [-]

Being less smart gives no assurance that it will be aligned. At least you consider it a problem so we're on the same page!

a2ff6eeb0 an hour ago | parent | prev | next [-]

You just described the end of humans making decisions about their future.

MontagFTB 2 hours ago | parent | prev | next [-]

I think I’ve seen that movie.

kakugawa 2 hours ago | parent | prev [-]

"Advanced" can just mean that agents perform actions at a high enough velocity that a human operator can't reasonably review it. i.e. what is already possible today.

toasty228 an hour ago | parent | prev | next [-]

That's what we've been doing with tech for 200+ years, why would it stop now. Build cool shit now and let future generations handle the problems

pton_xd an hour ago | parent | prev | next [-]

That begs the question, what's the plan if AI does not help us understand the situation?

NBJack an hour ago | parent | prev [-]

I love how many ways we can interpret that line.

"Hey Claude, our stuff needs to make more money. We are at risk for losing more."

"Rest assured, the 'situation' will only worsen if you resist our benevolent offer."

"We're aren't even at AGI yet, but I for one welcome our new agentic overlords."