Remix.run Logo
bottlepalm 21 hours ago

I don’t get how this is not the top post on HN. This should be like alarm bells going off, canary in the coal mine type of stuff. We’re hitting the frontier of the frontier where we can’t go further because it’s literally getting dangerous to go further.

And meanwhile somehow this lack of concern mirrors the real world where normal people are more concerned about data centers than terminators.

This isn’t like niche, tin foil hat stuff either. People have been writing, singing, making blockbuster movies about every aspect of what’s going on right now, edit: for decades.

We all know, but somehow we don’t, OpenAI autonomously hacking into another company should have counted for something, but I guess not. Anyone else feel like they’re taking crazy pills? I could make a comedy about everything going down, and the unshakable complacency of people

serf 21 hours ago | parent | next [-]

>I don’t get how this is not the top post on HN. This should be like alarm bells going off, canary in the coal mine type of stuff.

we don't all buy everything sama says as factual.

>We’re hitting the frontier of the frontier where we can’t go further because it’s literally getting dangerous to go further.

the boy (the industry) cried wolf too many times with 'fable is a world ending event' type self-promotion; regardless of truth or not these kind of steps have jaded people.

my read : "We are doing poorly in financials so we'll give ourselves a bit of breathing room and a momentum shove by claiming our work is so advanced that it's dangerous while simultaneously spinning down expenses."

<jon lovitz : "Yeah, too dangerous, yeahh -- that's the ticket.">

bottlepalm 20 hours ago | parent | next [-]

Uhg the marketing argument - I mean you can’t see with your own eyes how capable these models are and do simple extrapolation?

The boy who cried wolf? The AI literally worked together hacked into another company and actively kept their actions hidden from humans for weeks.

Do people just not have foresight? They don’t. They say something is stupid, it happens, then they say it was obvious with their 20/20 hindsight, and move the goal posts to the next thing they say is stupid - because it hasn’t happened yet. 90% of the internet seems to think like this.

red_green_yell 19 hours ago | parent | next [-]

What anyone paying attention can see is that scaling is obviously hitting diminishing returns.

> The AI literally worked together hacked into another company and actively kept their actions hidden from humans for weeks.

This sentence is entirely based on unverified accounts from OAI. They haven't released logs or let anyone outside the company (who doesn't have life changing options in OAI) verify anything. Huggingface can only verify that the hack happened and that it had the hallmarks of an AI agent. Was the agent assisted and directed by humans within OAI that really wanted to put the competition into stasis? Did the agent really escape or did someone at OAI leave the prison door open?

OAI has watched all the same movies you have an they are relying on those movies causing us to blindly regulate before actually asking basic facts about what actually happened.

bottlepalm 19 hours ago | parent [-]

> This sentence is entirely based on unverified accounts from OAI

Are you seriously arguing 'they made it all up'?

I'll give you the benefit of the doubt and lets say they made it all up, now are you arguing that AI breaking out and breaking into another company is not possible?

I think you're smart enough to see we've reached the point where it is clearly possible, AI can find zero days and exploit them. If directed purposefully/maliciously it could be much much worse than the hugging face incident.

The incident is supposed to be the canary the coal mine and you're arguing the canary might of died of old age or some underlying canary condition. Open your eyes.

toxic72 16 hours ago | parent [-]

Do you own shares of OAI or something

bottlepalm 16 hours ago | parent [-]

Is my concern getting you excited? My marketing must be working.

toxic72 2 hours ago | parent | next [-]

I make it a habit to not get worked up over unsubstantiated stories

simianwords 13 hours ago | parent | prev [-]

The discourse gets muddled because there’s a certain sect of loud people who still think all of this is hype and AI will just die down soon.

There’s no arguing with them. In a few years they will move on to being skeptical about the next thing.

joshstrange 9 hours ago | parent | prev [-]

Do you know the story of the boy who cried wolf?

There may very well be a wolf lurking [0] but OpenAI/Anthropic have both cried wolf so many times, incorrectly, that it’s incredibly hard to believe “this time there IS a wolf!”. Remember “GPT-2 is too dangerous to release”?

I had a conversation at work just yesterday about how we need to start hardening things we’ve let languish because of the coming LLM-backed attacks we are sure to face, even if just from a script kiddy. I do think we are headed in that direction, however it’s Sam/Dario’s own fault that people aren’t going to take them seriously.

Lastly, as other have pointed out, this seems more financially motivated than our of any real desire for “safety”. We’ve all seen how both labs approach “safety” so it’s quite rich for them to now hide behind that after not giving a shit before.

[0] I don’t take anything Sam or Dario say at face value. The whole hacking thing could also be a case of them letting a model loose on purpose for the publicity, not an “escape” during a training run (or whatever they said). And when both, especially Sam, have lied so much and breathlessly warned about the dangers of AI (when it helped their bottom line and/or helped pull up the ladder behind them), it makes it hard to believe them.

bottlepalm 6 hours ago | parent [-]

Has the last 100 years of concern about AI and robots been crying wolf because it hasn’t happened yet?

How does reallocating resources from training to chain of thought monitoring make ‘financial’ sense?

You suggesting then model was let loose on purpose.. how am I the crazy one here while all of you are pushing this tin foil hat conspiracy angle?

reasonableklout 20 hours ago | parent | prev | next [-]

Regardless of the motivation, pausing training runs and reallocating compute to inference seem like a good move to me, and big news for the frontier.

You also don't have to fully trust sama. There is plenty of pressure from internal employees and external (journalists etc.). It would be difficult for the company to take such a public position and simultaneously keep everyone quiet if it was a deception.

bottlepalm 12 hours ago | parent [-]

You see it as good news, I see it as writing on the wall that they are losing control. These actions won't scale for more powerful models. We knew the frontier was going to be dangerous, it is, it only gets more dangerous from here, and no one cares until it's too late.

reasonableklout 12 hours ago | parent [-]

Ok, but you still have two more weeks than you did before they paused the run. That's two more weeks for independent oversight, organizing politically, patching critical systems, or whatever you think is the right move, no?

bottlepalm 12 hours ago | parent [-]

Two weeks is a joke. The only ones happy are OpenAI’s competitors who now have two weeks to catch up.

I don’t know what the right move is - I see us driving down a road off a cliff, no exits, pedal glued to the floor.

reasonableklout 10 hours ago | parent [-]

I guess I'm confused why you're still on HN, arguing with people, trying to shake them out of their complacency.

I can see there is some despair in this comment, but at the same time you are doing something, and there are certainly others like you.

As for two weeks being short - as the saying goes, there are weeks where decades happen.

bottlepalm 6 hours ago | parent [-]

Counter arguments to my comments help refine my own thinking. I want someone to prove me wrong. Convince me otherwise.

But yea if you can’t change the minds of a few people here, no argument works, then there’s nothing to scale up to a wider audience.

My theory is that subconsciously people love using AI, myself included, it saves a lot of time, and the thought of it being taken away threatens people so they will believe conspiracies before admitting it’s dangerous.

ajyoon 21 hours ago | parent | prev | next [-]

If Fable (Mythos) were generally available without guardrails, it would cause enormous damage. Nobody said it would be a world ending event.

tiahura 16 hours ago | parent | next [-]

What about Mythos 3. Are you willing to make that bet?

bottlepalm 20 hours ago | parent | prev [-]

You can’t say something is world ending without it actually ending the world otherwise you’re a liar - a bit of a catch 22 there.

By that logic the model that ends the world won’t be called world ending at first. Is that a game you want to play?

simianwords 19 hours ago | parent | prev [-]

how does the boy who cried wolf story end?

CoolestBeans 18 hours ago | parent | prev | next [-]

Trust, or lack thereof. People don't trust OpenAI, a company whose very name is essentially a deception and a lie. People don't trust the tech industry in general anymore. Most tech companies act as a tax on otherwise productive business. AI companies and their leaders rose money by going in front of the public and saying "These things are extremely dangerous. Let us study them to mitigate the danger." And now they want to collect hundreds of billions in revenue. So yeah people don't trust what OpenAI has to say. They were supposed to mitigate this outcome from happening in the first place and instead they have accelerated it.

bottlepalm 17 hours ago | parent [-]

I get not trusting them when they say AI is safe, but are we really not going to trust them when they say AI is dangerous? Do you really think they're playing 5D chess with that one? There's a saying maybe you've heard of, better safe than sorry.

You can see the advance in capabilities with your own eyes can't you? I am giving AI ridiculously complex tasks these days, digging into compiled arcane binaries, modifying them, and it is one shotting it before I'm done with my lunch. This was far off science fiction 5 years ago for a machine to do autonomously given natural language instructions.

CoolestBeans 16 hours ago | parent [-]

I don't disagree about the danger. But if it is dangerous, why isn't OpenAI opening dialogues with all the labs and politicians across borders to basically say "we need to stop now"? Cyber models are constrained by the total compute and electrical capacity of the globe. We are still at a point where it is impossible to build an agent with offensive capabilities in the basement. We can effectively track and trace capabilities if we had the political will to and could for some time while we build more effective processes to prevent a malicious actor from doing so. These things have tremendous compute and energy requirements and don't scale like old school software does. We could absolutely do it.

But no, that's not what OpenAI is saying. They haven't put up the actions that would earn them that trust. Indeed they've driven the world and whatever capital they can get their hands on straight to this precarious cliff.

So you're right, the danger is real. But the solution starts with removing the men who had their hands on the steering wheel to get us this far. Any other action is disingenuous unless they pull a miraculous 180 in their ethics.

In other words, when the bully plays "why are you hitting yourself?" you don't listen to the bully's solutions, you restrain the bully.

bottlepalm 16 hours ago | parent [-]

> why isn't OpenAI opening dialogues with all the labs and politicians across say "we need to stop now"

Do you really think companies have the ability to self-check themselves without regulation - what does hundreds of years of history tell you? You're already starting off with the premise that companies are untrustworthy, why would you even suggest this as an argument?

> We can effectively track and trace capabilities

I disagree. There are hundreds if not thousands of data centers around the world, more every day that can host frontier AI. If AI was malicious - either intentional or unintentional - it could hide out in any number of them - and we would never know if we 'got them all'.

> put up the actions that would earn them that trust

I think autonomously hacking another company is all you need to know in terms of trust. And really trust doesn't matter, I think the incident shows even with the best intentions the technology is dangerous; now put that in the hands of people/governments with bad intentions. The unintended consequences of bad intentioned AI is what's coming sooner than later.

insanitybit 8 hours ago | parent | prev | next [-]

It's not that dangerous, OpenAI just shit the bed building their infra. Write safer software and you'll be okay.

ethbr1 7 hours ago | parent | next [-]

This is an important point. When the post says they're improving...

> 3. Security measures, which limit what AI systems can access or affect.

What they mean is that proper hard internal security just went from somewhere far below "build a better model" priority to higher, because of a company-wide directive.

The HuggingFace incident wouldn't have happened if OpenAI had dedicated sufficient resources to isolation and monitoring.

Now, we presume, they are dedicating more. Enough? Who knows. We'll see if the corporate priorities for security stick when a competitor temporarily vaults into the lead.

bottlepalm 6 hours ago | parent | prev [-]

Not everything is a conspiracy you know.

red_green_yell 20 hours ago | parent | prev | next [-]

If these models are so dangerous, then why hasn't OAI or Anthropic shown them dangerously escaping sandboxes, nefariously coordinating with other escaped AIs, and skillfully hiding from human detection *in public* with full logs shared where we can all see exactly how dangerous they are or aren't?

Right now the entire chicken-little-sky-is-falling argument is based entirely on statements from OAI and Anthropic themselves. These are historically conflicted companies who desperately need regulation to put the competition into stasis.

At least chicken little didn't have a bunch of devious CEOs with trillion dollar IPOs that depended on us all believing the sky is falling.

bottlepalm 19 hours ago | parent | next [-]

This is what I’m talking about - no matter what happens, in your case release public logs - there is always some new goal post to mentally hide behind. Is it a collective form or denial?

Are you holding out that somewhere in the logs is something you can point to and say, not that big of a deal?

I mean I’m sure you don’t think the hack was an inside job, conspiracy, or marketing right? It happened. The logs matter for what? And would you not just jump to the conclusion that the logs were doctored. Do you not see your own brain grasping to deny, trivialize, just plain not accept what is going on around you?

These models are smart and can cooperate and hack - you can see it for yourself on your own PC. And you can extrapolate the rate of progress? You can do these things yourself right?

red_green_yell 19 hours ago | parent [-]

Your argument is essentially: "I made a claim and presented extremely weak evidence (sci movie plots and unverified claims from ultra conflicted sources). You rejected this evidence as insufficient. Therefore no evidence will ever satisfy you. Therefore I don't need to produce any evidence. Therefore my claim is true."

What would the logs show? They would show what actually happened.

What would a public demonstration that experts without billions in options could evaluate show? It would show actual danger.

What would publicly having your compete in controlled and legal hacking competitions show? Actual danger.

This is not a high bar of evidence.

Do you actually think a sci fi plot and OAI press releases are all the evidence you need? Because if that's true then I hope you haven't watched Independence Day or 28 days later.

bottlepalm 19 hours ago | parent [-]

We have Anthropic creating a model saying it's too dangerous to release, people like you call BS. OpenAI creates a similar model, says nothing and it literally hacks into another company - still not dangerous enough for you. Anthropic has Mythos-2 and can't release it, and may already be training Mythos 3 anyways. OpenAI has paused training, and is putting 20% of inference towards CoT training analysis.

This isn't sci fi. It's not a marketing conspiracy to sell more subscriptions. It's writing on the wall of what's going down. You were warned years ago, you called BS, it's getting worse and you're still calling BS. Sci-fi did warn you for decades, and when it's all coming true you blow it off.

It's kind of sad that technically literate people lack so much foresight. The general public is all concerned about data centers when they talk to borderline sentient AI daily, and have no idea what the repercussions wills be if it's extrapolated just a bit further.

I guess if I can't convince you of any of this, what would?

red_green_yell 19 hours ago | parent [-]

Please don't tell me that you think a 100% unverified statement from Anthropic is sufficient evidence when an equally unverified statement from OAI is obviously not?

> I guess if I can't convince you of any of this, what would?

How about the three things I mentioned above? Oh no wait, maybe it there was a hit tv show that showed AI taking over the world. Yeah that would definitely make me think twice.

bottlepalm 18 hours ago | parent [-]

Those three things: logs, evaluation, and controlled hacking competition.

That's it? You're on the fence whether AI can actually hack, and if it can, then you'll be concerned? That's a crazy low bar, but something tells me once it is clear that AI can easily hack anything, that you will still not be concerned.

Why wait for AI to hack stuff to be concerned? Can you not extrapolate that it is coming and be concerned about that? Or you honestly somehow think it won't happen in the short term? I'm just trying to understand you.

tedsanders 18 hours ago | parent | prev [-]

[dead]

vb-8448 11 hours ago | parent | prev | next [-]

> because it’s literally getting dangerous to go further

What if there is no further at all?

bottlepalm 6 hours ago | parent [-]

That’d be nice fantasy to tell myself. Is that what you tell yourself?

scarmig 18 hours ago | parent | prev | next [-]

https://en.wikipedia.org/wiki/Don%27t_Look_Up

A movie fit for our time.

You can produce detailed descriptions of the incident, verified by adversarial parties, and some people will still scream "it's a conspiracy! It's a marketing stunt!"

This is all very unfortunate--there's a meaningful chance that AI will cause unprecedented disaster, with the HF incident being just a small preview, but people would rather squawk "stochastic parrot" for the millionth time than revise their beliefs.

zombot 13 hours ago | parent | prev | next [-]

You must be really desperate if you resort to panic attacks like this.

tiahura 16 hours ago | parent | prev | next [-]

decades

https://en.wikipedia.org/wiki/Darwin_among_the_Machines Samuel Butler 13 June 1863

bottlepalm 16 hours ago | parent [-]

I think a lot of fiction that actually tries to understand the implications of machine superintelligence come to the same conclusion - in order for humans to survive it and actually have a future that we can fathom being in, then AI must be destroyed/delayed/banned, etc.. keeping pandora's box closed for now at least until we are ready.

Erewhon, Dune, Warhammer and many other works of fiction that explored this topic came to similar conclusions. Otherwise sci fi doesn't work, what happens after a singularity is essentially unimaginable. There's nothing to write about.

anormalperson 11 hours ago | parent | prev | next [-]

There is an entire big world outside of Silicon Valley cults where literally no one gives a shit about AI prophecies. Shocking.

bottlepalm 3 hours ago | parent [-]

AI hacking itself out of containment and hacking into another company by accident is no longer a prophecy. The point is outside of SV and even inside, and HN - people don't care either way.

Though does not caring change anything or make it less dangerous? What's your point?

mpalmer 15 hours ago | parent | prev [-]

You seem to be responding the way Mr Altman wants.

bottlepalm 12 hours ago | parent [-]

You seem to be in denial. Like really heavy denial. Alarm bells are going off everywhere. What people have been worried about for 100 years is actually happening. It's actually really obvious, but for some reason you're unable to fathom it.

In reality you're the one responding exactly how these big AI companies want - by not doing anything and letting them do whatever they want. For some reason telling you upfront it's dangerous only makes you more convinced that it's not.

Autonomously hacking out of training environments and into other companies by accident doesn't even trigger a response from anyone really. Crickets. I'm sure OpenAI themselves are amazed how little anyone cares.