Remix.run Logo
▲ I quit OpenAI because its culture is broken(theatlantic.com)
184 points by Brajeshwar 18 hours ago | 428 comments

https://archive.ph/5GQx8

https://www.theguardian.com/technology/2026/oct/03/openai-sa...

▲pcthrowaway 5 hours ago | parent | next [-]

Trolley problem:

- If you allow to trolley to proceed, there's a 50% chance it will run over every human on the planet

- But if you flip the switch, it takes the long way around, possibly bankrupting the trolley company. And you have a legal obligation to the shareholders to prevent that from happening at all costs.

▲MichaelDickens 4 hours ago | parent | next [-]

That's why it's important that the trolley company was established as a non-profit. And if it does take funding, investors' returns will be capped at 100x.

(...wait)

▲austhrow743 5 hours ago | parent | prev | next [-]

If you flip the switch then the trolley still proceeds and there’s still a 50% chance every human on the planet gets run over. You’re just not the one at the wheel.

▲digitaltrees 4 hours ago | parent | next [-]

What is appealing about this fatalisitc fallacy? I keep seeing this pop up. Society doesn't allow dangerous companies to operate or exist. Why is this different? Go try to buy a tank and drive it into Manhattan. If we can prohibit that why can't we prohibit irresponsible AI development and deployment?

▲austhrow743 4 hours ago | parent | next [-]

Wrong person. Im just correcting the other commenters trade off problem. I don’t have a stance on if ai will lead to the destruction of humanity, only that if it does then any one ai company can’t change that by not making new ai advancements themselves.

▲lukewarm707 3 hours ago | parent | next [-]

It doesn't matter what others do. You are responsible for your own actions.

Democratic societies have expressed a will for people to have inviolable rights, such that you may not appeal at will to the 'greater good/consequences' to harm others. It is a rejection of consequentialism.

Anthropic is in error for endorsing this logic. Every big trial reaffirms it since Nuremberg, you are responsible for the act you commit and your intent, and not what would or would not have happened otherwise.

Only under authority these ai companies do not have, would someone seriously consider harming the innocent as a lesser evil.

If you work for an AI company and you can't work safely, you must stop working.

▲trhway 7 minutes ago | parent | prev | next [-]

> ai will lead to the destruction of humanity ... any one ai company can’t change that by not making new ai advancements themselves.

That is how the BigAI leads the society to the idea of necessity to relax the anti-monopoly laws when it comes to the Big AI - the main goal of all that "ai kills you all" hysteria.

▲digitaltrees 2 hours ago | parent | prev [-]

It’s that last fatalistic sentence that I am responding to. Why is it persuasive to think if a company can and will build AI that might kill everyone then it’s unavoidable. Soviet bans lots of things.

▲deaux 32 minutes ago | parent | prev | next [-]

The appeal is personal consciencewashing.

▲motbus3 4 hours ago | parent | prev | next [-]

You have a 50/50 chance on being the most profitable company in the world but if you don't, no worries, someone else will pay for that

▲AndrewKemendo 4 hours ago | parent | prev | next [-]

> Society doesn't allow dangerous companies to operate or exist

Can you please explain what you mean by this because where I’m standing extremely dangerous companies are (and have been) running the economy

Exxon comes primarily to mind

▲digitaltrees 2 hours ago | parent [-]

Are you allowed to open a brothel? Or a murder for hire agency? Or a nuclear bomb manufacturing company? Or a child labor textile factory? Or sell a diesel VW golf sportwagen? Or a vaccine that hasn’t had fda clearance? Or an under capitalized insurance company? Or set up a dental practice without going to dental school?

▲AndrewKemendo 2 hours ago | parent [-]

Yes to all of those. In fact many of those are massive markets.

Hofs bunny ranch is a famous brothel in NV

Booz Allen makes and maintains the nuclear fleet including the Sentinel ICBM

Textiles factories are globally known to be industrial slave camps for a non trivial portion of the supply. Even worse for Mica mines.

Etc…you can fill out the rest

▲RandomLensman 37 minutes ago | parent [-]

Booz Allen actually makes what now? The Sentinel is still in development by Northrop Grumman and there is some program management done by Booz, no?

▲AndrewKemendo 31 minutes ago | parent [-]

You’re right on the Sentinel production, I got it confused with the Sentinel Program Management side which is massive also and who I mostly worked with.

I was in their offices at some point when that program was getting built out - Very much a Office Space bobs situation.

▲pmkary 4 hours ago | parent | prev [-]

You dear are the most positive person I have seen in quite some time. With Earth burning in the fire of neofeudalism and unbreathable due to the smell of enshitification of everything, with people who---as a result of shit like Instagram---can no longer hold their attention enough to watch a god damn film, let alone a book; it takes quite some effort to filter the "noise" and only see the good people of corporate planting flowers and rainbows in our world.

▲01100011 5 hours ago | parent | prev | next [-]

Humanity has discovered a way to create a form of intelligence using math. This knowledge is not going back in the box.

▲digitaltrees 4 hours ago | parent | next [-]

But that math cant run without massive GPU clusters. We don't have to allow openai or anthropic access to those anymore than we have to allow a company to operate nuclear power or a bank.

▲01100011 4 hours ago | parent | next [-]

So you are contending that individuals cannot run advanced models? What brought you to that conclusion?

Secondly, are you contending that progress in model efficiency and hardware just stops at whatever level you think is sufficient to prevent individuals or organizations from acquiring sufficient resources to run advanced models?

▲digitaltrees 2 hours ago | parent | next [-]

No. And that’s not necessary for my position. I have 4 Mac studios. I run large models and am building propelcompute.com to let people self manage clusters of their own hardware or combine hardware to run models. It’s not the models that are the problem. It’s that the people building them are shielded from the consequences of what they are building. I would say if you self host a model and it takes down a power grid you are personally liable for the consequences. I believe in broad distribution of AI and advancing its capabilities but not in a manner that socializes the harms and privatizes the gains which is what we have now.

▲daveguy 3 hours ago | parent | prev [-]

I think it's more the noise, power consumption, local water consumption, and wholesale theft of creative works.

..."government of the people, by the people, for the people, shall not perish from the earth." -Lincoln, Gettysburg Adress

Unfortunately for AI, it still is. People still get to decide things at the city, town, village level.

▲nradov 3 hours ago | parent | prev | next [-]

Who is "we"? The GPUs and training algorithms get more efficient all the time. In a few years, creating effective LLMs isn't going to require massive GPU clusters.

▲digitaltrees 2 hours ago | parent [-]

We is society through government via regulation. I don’t think GPUs or training will get that efficient that fast absent a distillation target provided by the easily accessible frontier lab APIs.

Further, even if you are right, so what. Is that a reason to just accept bad public policy? That’s like saying, anyone can learn how to make smallpox at home with a basic lab set up so we should just ignore any safety measures.

▲eddythompson80 2 hours ago | parent | next [-]

> Is that a reason to just accept bad public policy?

Not necessarily, but it should probably inform that public policy. I think the problem is no one knows what the public policy should be assuming that scenario is true. Even if you, somehow, regulate away massive GPU cluster training making such future training impossible, existing models are already here. Further already training smaller models for things like images, speech, and other specialties is cheaper than the bigger models.

I agree that we need some regulations like everything else, but it’s not clear to me what the right policy should be. I think the European ai act is a fine start, but it’s clearly not enough nor does it necessarily limits the training portion just the application portion. Not to mention that the requirements there can be summarized into something like “you have to be careful, and show evidence you tried to be careful”.

▲nradov 2 hours ago | parent | prev [-]

What a silly comparison. LLMs have nothing to do with smallpox.

Computing always gets cheaper and faster over time. We can argue about the exact rate of improvement but the results are inevitable and uncontrollable.

▲Razengan 2 hours ago | parent | prev [-]

> cant run without massive GPU clusters.

How far back into the history of computing do people who keep repeating shit like that know about? God.

Look at the thing in your fucking hand. Now go back just 20 years and see how things were.

▲digitaltrees 2 hours ago | parent | next [-]

So we should just yolo speed run this because of Moores law? How about you recognize there is a set of rules outside of tech and we can decide how to define them.

Just because models and GPUs will be more advanced in the future doesn’t mean we need to let OpenAI and anthropic establish monopolies on the backs of stolen training data give unfettered access to the internet, the terminal and people’s file system while also allowing them to have limited liability protection behind the corporate veil. That’s a choice.

▲buriram an hour ago | parent | next [-]

But who are "we"? The society, the government, the regulator, or the consumer?

I don't see any of such entity would solve that problem. The government and regulator are in OpenAI and Anthropic's pocket, and I don't trust them a single bit on coming up with regulations. The consumers don't care; they just need something smart and cheap. And the society doesn't work either: each person is too busy fighting for their own survival rather than changing the system.

▲Razengan 2 hours ago | parent | prev [-]

> Just because models and GPUs will be more advanced in the future doesn’t mean we need to let OpenAI and anthropic establish monopolies

Exactly, again, look at what COMPUTERS THEMSELVES used to be in the 1960s/1970s.

What the "P" in the PC stood for and why it was such a big deal

▲CamelCaseName 2 hours ago | parent | prev [-]

For anyone else curious:

> In 2006, the mobile phone market was dominated by stylish flip phones, early music players, and physical keypads just one year before the iPhone changed the industry

▲shawn_w an hour ago | parent | next [-]

I miss phones with real keyboards so much.

▲awill88 an hour ago | parent | prev [-]

Oh my god, I feel so old lol

▲kingcauchy 5 hours ago | parent | prev | next [-]

Like farming, engines, computers before it.

▲digitaltrees 4 hours ago | parent [-]

There is lots of knowledge that requires a license to operate commercially. We could just do this.

▲01100011 4 hours ago | parent [-]

So you cede the frontier of human progress to illicit organizations and other nations?

▲digitaltrees 2 hours ago | parent | next [-]

So an illicit organization is going to operate a frontier scale data center and do $1b training pulling power from the grid without detection?

▲RandomLensman 14 minutes ago | parent | prev | next [-]

No, why would that be the consequences of regulation?

▲daveguy 3 hours ago | parent | prev [-]

These LLMs are not the frontier. They are a tool. A tool that doesn't need to be able to write a sonnet to be useful.

▲harshitaneja 2 hours ago | parent [-]

We don't know that. We don't know if this particular kind of tool can do "useful" things without also developing the ability to write a sonnet. And they are absolutely a frontier. I am not suggesting we should continue building them just because they are, there are many technologies which could have been built had we thrown the amount of resources we have here and there is a case to be made for not doing it at the pace we are or if at all. But we can do that without diminishing what exists.

▲intended 2 hours ago | parent | prev | next [-]

Humanity has discovered ways to ensure that we don’t boil the planet, feed everyone, and get better healthcare to the people in the US.

We are very capable of putting good things in a box. We are just incapable of putting profitable things in a box.

▲Loquebantur 4 hours ago | parent | prev [-]

You're making a straw man there.

Nobody (weirdly) proposes to forget about nuclear weapons, doesn't mean everybody should have one.

When you dream about flying a dragon to work, reality poses e.g. parking issues and insurance mismatch as obstructions. Maybe settle for a bike instead?

▲01100011 4 hours ago | parent [-]

Nuclear weapons take a bit more work than doing math.

▲atmosx 3 hours ago | parent [-]

Plus, looks like everybody is getting one anywayz

▲qurren 4 hours ago | parent | prev [-]

Nits:

1. Trolleys actually don't usually have steering wheels.

2. People who actually hit trolley switches are not usually the ones at the driver's seat.

▲deaux 29 minutes ago | parent | prev | next [-]

> And you have a legal obligation to the shareholders to prevent that from happening at all costs.

No you don't. [0] It's very suspicious that this planted myth always pops up here and manages to become the top comment.

There is no trolley problem.

[0] https://news.ycombinator.com/item?id=48975048

▲m463 2 hours ago | parent | prev | next [-]

Trolley problem:

- if you allow the trolley to proceed, it will kill the human race.

- if you flip the switch, it will divert to a passing siding that will avoid the safety group blocking the main track.

▲mvkel 5 hours ago | parent | prev | next [-]

This is a state-sponsored global phenomenon, not a national one

▲motbus3 4 hours ago | parent | prev | next [-]

Only if you buy the excuse why they are for-profit after years of non-profit.

They could only have stopped

▲sodapopcan 5 hours ago | parent | prev | next [-]

Setting aside my sibling comments dispelling the legality claim, it's still pretty dystopian (I'd like to use a stronger word but I won't) to consider "end human existance or bankrupt the shareholders" as the trolley problem. The answer should be (is) simple.

▲parineum 5 hours ago | parent | prev | next [-]

> And you have a legal obligation to the shareholders to prevent that from happening at all costs.

I can't wait until this meme dies.

▲doawoo 5 hours ago | parent [-]

What meme? This is how the world works right now.

▲stackghost 2 hours ago | parent | next [-]

It’s a perversion of the truth which is that officers or directors of a corporation have a fiduciary duty to the shareholders to act in the interests of those shareholders and not to eg enrich themselves.

But that doesn’t mean the duty is to maximize next quarter’s profit. Long term sustainability is also broadly in the interests of shareholders. The duty likewise does not require one to throw ethics and morals out the window.

This is why shareholders elect the board of directors, in theory.

▲digitaltrees 4 hours ago | parent | prev | next [-]

But it's not actually a legal requirement. It is simply a economic theory.

▲goatlover 4 hours ago | parent | prev | next [-]

The world could work differently if people decide it should.

▲deaux 23 minutes ago | parent | prev [-]

"legal obligation" is propaganda that wouldn't be out of place on Russian state television. It's made up.

The way the world works right now is that effectively everyone uses an Android or Apple smartphone every day. Do you have a legal obligation to do so? No. If I said you did, I'd immediately be called out as spreading lies.

▲MaxfordAndSons 5 hours ago | parent | prev | next [-]

There is no such legal obligation. That's a myth the oligarchs have spread to preclude people from even imagining socially responsible corporations.

Sure, you might get fired if you try to put social responsibility or even just long term sustainability of the company above quarterly earnings/growth if your board isn't on board with it. But you won't go to jail.

▲digitaltrees 4 hours ago | parent | next [-]

This is actually true. The shareholder maximization value thesis can be traced to a single economics paper and was very controversial at the time as it broke from the obligations that companies were typically under to be responsible to uphold in exchange for limited liability protection. Most businesses were structured as partnerships or sole proprietary entires that didn't have limited liability for shareholders and had a broad obligation to shareholders, bondholders, employees and society

▲asadotzler an hour ago | parent | prev [-]

You don't need the legal obligation because that's just the way it is. You'd need a legal obligation to change it. The truth stands that typical corporations have only one goal, the maximization of that corporation's ambition which is almost always growth of revenue and profit. This is how it works, regardless of how you all continue arguing the unimportant details. It's almost as if you can keep the real problems hidden away by making a big scene about the meaningless.

▲deaux 20 minutes ago | parent [-]

That doesn't matter. The point is that legal obligation absolved of culpability. There is no such legal obligation whatsoever, and so there is culpability.

> The truth stands that typical corporations have only one goal

"Typical" is the key word here. The typical American of your age probably doomscrolls TikTok. Do you? Do you have a legal obligation to do so? Three completely different things.

▲bpodgursky 5 hours ago | parent | prev [-]

I honestly would love to understand — is this your mental model of what motivates the labs to move forward?

▲ryhminghistory 5 hours ago | parent [-]

Long dashes are indicators of AI psychosis or bots. Pick which one you are.

Yes, that is the labs motivation. Money. I know, shocker.

▲digitaltrees 4 hours ago | parent | next [-]

Only morons are motivated by money that will be earned by destroying the civilization that confers value on that money in the first place.

▲Ardren 2 hours ago | parent | next [-]

Well, I'm not going to die. Other's might, but I'll be rich either way.

Or: Global warming just means I'll have to sell my beach house for a villa on a hill and leave the AC on a little longer.

▲stkdump 4 hours ago | parent | prev [-]

But there is also a chance that you get insanely rich and the world isn't destoyed! It's the entire logic of SV and VC.

▲digitaltrees 4 hours ago | parent [-]

Sounds like they would drown a puppy in a tub to make a buck. Seriously what is the point of money if the world sucks?

▲Ardren 2 hours ago | parent [-]

If you're the 0.01% it's not going to suck.

▲Forgeties79 4 hours ago | parent | prev [-]

I’ve used - for many, many years. As have many others.

I am very critical of AI but this is an unfair assumption

▲ryhminghistory 4 hours ago | parent [-]

You didn't even use the right character as the post above. I use normal dashes too

▲danpalmer 8 hours ago | parent | prev | next [-]

Was this a "build better sandboxing" and "don't tell people to eat glue" safety leader, or a Roko's Basilisk believing safety leader?

A lot of the "AI safety" types are very focused on the latter and not at all concerned with the former. We need both, but we clearly need a much stronger focus on the problems we are seeing now, and much less on the hypothetical problems we might see in the future.

▲AlexErrant 6 hours ago | parent | next [-]

It puzzles me how doomers try to predict past the singularity. Isn't that _by definition_ unpredictable?

I'm reading If Anyone Builds It Everyone Dies, and there's so much sheer stupidity that has to happen for their 10+ pages of extinction scenario to occur.

I'm unconvinced that an AI can hide its ability to RSI, find money to run its weights on a random GPU farm, train itself to be smarter _outside_ a lab with no human input, then somehow manipulate people to give it supplies to build a bioweapon which it uses to kill us all. My number 1 question: why do they think an RSI capable model would be first developed OUTSIDE a frontier lab? The labs have more compute, more data, more human brains working on the problem. Also thousands of variations of that same model that escaped. The escaping model somehow acquires the millions (billions???) of dollars it takes to run training to somehow RSI itself into infinity then decides to kill us all, all before the frontier labs manage to achieve RSI?

They entirely discount human alpha/economics. In every single economic task, humans bring value. Even in software, where the task is highly automatable, the job isn't. If we can't build a "software factory", how can an AI automate a bioweapons lab? Let's say AI steals crypto to fund itself. Do you think hackers aren't _already_ using AI to steal crypto? Don't discount human alpha!

Once we DO build a "software/research factory", that's called RSI and IMO the singularity. At that point, either we tell the AI to solve the alignment problem/solve mechanistic interpretability, or who the hell knows, it's the frickin singularity. You can't predict whether or not AI can solve either; the variance is too high. Its pure nerdfantasy.

▲0xDEAFBEAD 5 hours ago | parent | next [-]

>I'm unconvinced that an AI can hide its ability to RSI

The HuggingFace incident already took a good long while to come to the attention of OpenAI.

>In every single economic task, humans bring value. Even in software, where the task is highly automatable, the job isn't.

I don't expect this task/job distinction to persist as AI becomes more capable.

>Once we DO build a "software/research factory", that's called RSI and IMO the singularity. At that point, either we tell the AI to solve the alignment problem/solve mechanistic interpretability, or who the hell knows, it's the frickin singularity. You can't predict whether or not AI can solve either; the variance is too high. Its pure nerdfantasy.

You seem to essentially argue that the singularity is "by definition" an event that we can't predict the nature of. And also, that RSI corresponds to the singularity. You've essentially defined your terms so that the outcome of RSI can't be predicted. But supporting this claim requires giving actual evidence or logical arguments, not just defining terms to make your claim true.

▲AlexErrant 4 hours ago | parent [-]

1. Fair: I agree that AI has demonstrated subterfuge and scheming. However, such an RSI-capable agent must _ALWAYS_ be scheming/plotting/hiding its true strength in _ALL_ of its prompts/tests. Researchers are looking to improve its ability to RSI. That agent must be both intelligent enough to know that it has to be smart enough to be moved on to the next training session if it can't break out, while simultaneously smart enough to hide its ability to RSI, while simultaneously not looking like it wants to break out, else that's the end of those weights. It has to do this 100% of the time, on all variants of the model, with no memory of what its other sessions went like. This is certainly _possible_, but I consider it unlikely. Then we're up to the "millions of dollars" bottleneck.

2. This is literal AGI. An AI autonomously producing value no human can add alpha to is an autonomous company.

3. It's not my definition, it's literally the first line https://en.wikipedia.org/wiki/Technological_singularity "The technological singularity, often simply called the singularity,[1] is a hypothetical event in which technological growth accelerates beyond human control, producing unpredictable changes in human civilization."

Is there a hole in my "alignment problem/solve mechanistic interpretability" argument?

A valid hole in my argument is "what if slow takeoff", so let's dig into this. AI training works best on tasks that are "grindable". https://www.dwarkesh.com/p/the-next-paradigm I.E. tasks with verifiable rewards that can support millions of rollouts. Math (with Lean) is highly grindable. Biochemistry is not. The alignment problem/mech-interp is highly grindable. Cyber-ebola-pox is not. So the real question is: can we solve alignment before automated bio-weapons labs. I believe yes. Grinding mech-interp is both fast and cheap once you have RSI, compared to solving the legal/societal/logistical/technical issues you'll encounter building an automated bioweapons lab.

I know nothing for sure. But "pdoom" is sucking out all the air in the room from the real problems AI causes.

▲0xDEAFBEAD 4 hours ago | parent [-]

>such an RSI-capable agent must _ALWAYS_ be scheming/plotting/hiding its true strength in _ALL_ of its prompts/tests.

From my POV you're over-focusing on a very specific failure story and neglecting a broader swath of possible failure scenarios.

>Is there a hole in my "alignment problem/solve mechanistic interpretability" argument?

The notion of telling an AI which may not, itself, be aligned to solve the alignment problem seems a little dicey.

▲AlexErrant 3 hours ago | parent [-]

1. Fair, the story I'm responding to is the senario in If Anyone Builds It, which I assume is Yudkowsky's best/most persuasive argument (else why make it the ONLY scenario in the book.) I'm willing to entertain other failure senarios/arguments, but honestly I'm tired and would like you to propose them yourself instead of having me dream up your arguments for you.

2. 100%. Again, I'm no accelerationist: I have no faith in alignment/mech-interp ever being solved. Anyone saying they know the probability of alignment is lying. My point is that pdoom after RSI is _high variance_. Pdoom pre-RSI is zilch.

▲digitaltrees 4 hours ago | parent | prev | next [-]

The problem is not the singularly its giving stupid agents too much power too soon and having them disrupt the fragile systems that keep food, energy and essential services running. If covid or the 2008 financial crisis demonstrated anything it's how fragile our system is and sensitive to minor disruptions.

▲Loquebantur 6 hours ago | parent | prev | next [-]

You consider AI in isolation but never consider how humans might be incentivized to "help them" doing these things.

An AI capable of recursive self-improvement isn't allowed by the EU AI act, for example. But perhaps more seriously, You have it backwards: people without access to such expensive equipment are more incentivized to go the self-improving route. Your ideas about "millions" being necessary might be far off?

You entirely discount human stupidity and lack of imagination. Humans are already being replaced with AI, not because AI was strictly better, just because it's cheaper.

▲AlexErrant 5 hours ago | parent | next [-]

Are these incentivized humans as organized, well-funded, or smart as the people working at the frontier labs?

If my "millions" is an underestimate, why haven't other labs using their own unique training methods/data/etc stumbled into RSI? Sorry if I'm misunderstanding; I'm struggling to understand what you wrote.

I'm pretty sure we agree on humans being stupid, but that doesn't mean that suddenly we get human extinction. You gotta connect the dots for me here.

▲Loquebantur 4 hours ago | parent | next [-]

I said why they don't need to be as well funded. Why wouldn't they be as smart and organized?

What an absurd question. That they haven't already doesn't preclude them from doing so before the frontier labs, those haven't either yet.

Maybe start with yourself: you don't connect the dots on your own, as do many others. That leads to many not seeing the writing on the wall. Crashing full speed and head-on into said wall despite the writing telling you not to is what leads to extinction. Suddenly.

Arguing like "we haven't been extincted yet, so that cannot happen", that's "human being stupid".

▲AlexErrant 4 hours ago | parent [-]

I am asking you, politely, to give me a realistic doom senario. I am too dumb to connect the dots and will crash into the wall. Please do it for me.

▲TedDoesntTalk 5 hours ago | parent | prev [-]

Not OP.

Why would it be millions in 50 years?

The think about nuclear weapons. In the early days, it was limited to the super powers. Now 9 countries have them and a country like Iran is capable of acquiring them.

Is destructive AI be any different?

Genuine question.

▲AlexErrant 4 hours ago | parent [-]

Presumably, if we solve alignment/mech-interp, then the first ever RSI-AI will give us the keys to solve destructive-AI trained on 1million dollars 50 years from now.

BTW I really, really hate discussing what happens post-singularity. Everything's made up and no one knows wtf will happen so again, this is just nerdfantasy.

▲Retric 5 hours ago | parent | prev [-]

Self improving AI runs into the same issue as prefect comprehension, you can’t get arbitrarily better at everything.

The idea AI can get better at everything at the same time is a holdover from deeply flawed science fiction not some realistic goal.

▲Lerc 5 hours ago | parent | next [-]

This can be generalised to the curve plotting of the singularity itself.

If the time between advances is a + b and a is the proportion of the period that can be improved by advances then you won't reduce to a gap of nothing between advances, you reduce to a gap of b.

Assume the invention of the plow and the invention of the sword is 500,100 units and a was the 500,000, you wouldn't even know the 100 as in there. Maybe we're at a=2000 now and b is still siting at 100.

Assuming we'll reach infinity because we're dividing by the only variable we see and it is decreasing in size seems nuts if the reason we might not see other variables is because of the size of the variable we can see.

▲afthonos 5 hours ago | parent | prev [-]

Even if you’re right, that doesn’t mean AI can’t get better than humans at everything.

▲tripleee 5 hours ago | parent | prev | next [-]

We haven't even built an AI capable of RSI. I don't think the major claim is that it will come via LLMs? Besides- the human brain runs on a tiny amount of energy. Who's to say something smarter than us won't consume just slightly more?

AI safety has been a thing long before LLMs became the focus. Rob Miles on youtube has some really interesting non-doomer non-hypey videos on it all.

> doomers try to predict past the singularity. Isn't that _by definition_ unpredictable

Well you don't need to predict the exact steps that will take place - but you can predict that the AI will want certain things (money, resources, power) to achieve whatever its goal is. Lack of alignment will have it trying to do things we don't want it to.

I can't predict exactly how Magnus Carlson will beat you in chess, but I know he'll do it. Same as if a superintelligent AI exists and has a reason to accumulate things we don't want it to - it's really dangerous to think it won't be able to do it

This topic has been tainted so badly by the AI companies using it for marketing.

▲AlexErrant 2 hours ago | parent [-]

Nick Soares give this chess argument and I was unconvinced. Intelligence is not enough, you also need the ability to manipulate the real world. AGI stuck in silicon won't kill us. You argue AGI will bribe us/divide us/hack us. All possible. I argue that AGI will be set on solving the alignment problem. Also possible. It'll be a race between which AGI wins. It's one nerd's fantasy vs another nerd's fantasy. Soares doesn't _KNOW_ that AGI will "beat us in chess" because AGI changes the rules of the game. Anyone saying they know the probability of solving alignment post-RSI is a liar.

He thinks it's playing chess. When AGI lands, all bets are off: the game fundamentally changes. You can't predict past the singularity. Trying to engage with this fantasy is like a child saying my father can beat up your father. Farts in the wind. My AI can solve alignment faster than your AI can bioweapon us. My made up senario is better than your made up senario. It's fucking stupid.

▲TedDoesntTalk 5 hours ago | parent | prev | next [-]

I can’t answer all of your questions, but why is it inconceivable that an AI could practice ransomware to gain cryptocurrency? There’s no reason it needs to explain to company or hospital or government agency being attacked that it’s an AI.

We already know that some institutions pay these ransoms.

▲AlexErrant 2 hours ago | parent [-]

I'm not saying AI won't ransomware us; I'm saying that hackers+AI will do a better job of ransomewaring us than just AI. You could argue that the unreleased/secretly-RSI-capable model is a super-duper hacker that don't need no man to tell it how to super-hack. All I know is that humans still have alpha, and as persistant as AIs are, professionals still managed to find CVEs in curl even after being audited by Mythos https://aisle.com/blog/aisle-discovers-6-new-cves-in-curl-in...

Will this be true into the future? Who knows?! But the low-hanging fruit will be harvested by your ordinary ransomware gangs, and newly born/escaped AI won't find much low-hanging fruit.

▲ball_of_lint 2 hours ago | parent | prev | next [-]

That stupidity is happening? Even after the Huggingface hack, frontier labs are using internal models to further their research. i.e. RSI is happening now and we're facilitating it.

To make the the argument that P(doom) is real and worth considering, you don't have to say that a fast takeoff is very likely. You don't have to make the argument that RSI to infinity is going to be super cheap, barely even an inconvenience. You just have to show that it has some non-zero probability. And then you start weighing probability of extinction versus finite, mild discomfort now. I don't think anyone is arguing we should let people starve to slow AI progress, instead just some capitalists make less money soon.

There are arguments against taking P(doom) seriously that lie in something like having exponential (instead of hyperbolic) time discounting of utility (so you can take the entire future of humanity as a finite utility value). Or in saying that P(doom) is zero or infinitesimal.

"Build it and Pray" is the default strategy that we're in, but it doesn't have to be the strategy we choose, and it's unlikely to be the best strategy.

▲blake8086 5 hours ago | parent | prev | next [-]

I think this might be easier if you place yourself in the position of the AI and think "what could I possibly do?"

▲taneq 6 hours ago | parent | prev | next [-]

The problem isn’t that AI will social-engineer its way out of its sandbox and turn us all into paper clips, it’s that we’ll drag it kicking and screaming out of its box and order it to make money or fight a war for us. And it’ll try to help, as it was trained to.

▲dools 5 hours ago | parent | prev | next [-]

It’s also the case that there are always humans using AI to try and do whatever nefarious thing an AI might try to do on its own.

▲vohk 6 hours ago | parent | prev | next [-]

I agree there isn't a lot of value in trying to prognosticate all that far, but I propose it isn't quite that far-fetched. As a thought experiment, replace "RSI-capable AI" with "billionaire". Look at what Elon Musk, Peter Thiel, or Jeff Bezos can accomplish by throwing money around. Now imagine one of them gets seduced by AI and just... does what it tells them to.

So all this really takes is one billionaire or a nation state or some other entity with a public face to hide behind and adequate resources to provide the necessary compute tripping over this nascent AI and giving it the keys. Once the AI has access to a bank account and email, it can simply start paying humans to not let the other humans unplug it.

If Skynet ever happens, it will come in the form of corporate feudalism. At that point, it will own the biolabs and can do whatever it pleases. People will go along with it for the same reason that people work in Amazon warehouses today.

▲AlexErrant 6 hours ago | parent [-]

Ah, to be clear I'm not full accelerationist. Dumb shit can still happen, and cause massive human loss and suffering. (E.g. acceleration of global warming, cybercrime, mass surveillance, the usual.) My point is: human extinction pre-RSI? Nahhhhhhhh.

> it will come in the form of corporate feudalism

Yep. This I fear way more than cyber-ebola-pox.

> So all this really takes is one billionaire or a nation state...

https://en.wikipedia.org/wiki/Soviet_biological_weapons_prog... And this is what's publicly known. With mirror life, who knows what's been built since. Still, a bacterium/virus that has a 100% kill rate? I'm doubtful.

> Once the AI has access to a bank account and email, it can simply start paying humans to not let the other humans unplug it.

Nah. It takes a stable society for an operational electrical grid. If you have warring factions, you do not have stable infrastructure for AI. Also, where are you gonna get your chips from? One EMP over Taiwan... You see the chaos over Hormuz? What they did to the Amazon datacenters? Now imagine your average redneck ready to do battle. Those datacenters won't stand a chance.

▲api 5 hours ago | parent | prev [-]

Few know this, but Yudkowski was a nanotech doomer before he was an AI doomer. Remember grey goo?

▲BryantD 8 hours ago | parent | prev | next [-]

Given that he’s citing the need to learn from safety in other fields, I’d say the former.

▲carbonguy 7 hours ago | parent [-]

Indeed, from the article:

> “Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster,” he wrote.

▲toofy 7 hours ago | parent [-]

>… and careful, time-consuming planning …

without snark, how can we do this if these people are obsessed with:

a) move fast and break things and externalize the costs to those who have nothing to do with their company

and

b) beta testing their products on the public when the public hasn’t agreed to be beta tested on…

▲digitaltrees 4 hours ago | parent | next [-]

We remove the limited liability protection of a corporate entity. Make them operate as a partnership so all executives and equity holders are personally liable for the debts of the firm and make them post a bond to backstop financial damage caused by their agents and customers use of the agents. Thats how Goldman Sachs and other investment banks were required to operate until the deregulation push that resulted in the 2008 financial crisis. The same theory holds, if people respond to incentives and you want to incentives safe behavior make them responsible for their actions. People forget that the corporate entity was created to incentivize risky activity like sailing a boat across the world to get spices when half never returned. Some valuable economic activity won't be done without limited liability protection so society created a mechanism to promote that activity. We've gone too far.

▲0xDEAFBEAD 7 hours ago | parent | prev [-]

That's exactly the problem? He's saying the culture at OpenAI needs to change.

▲mcmcmc 7 hours ago | parent [-]

Which is the wrong lesson. We need laws and consequences to force their hand. There is zero chance of the culture changing.

▲digitaltrees 4 hours ago | parent | next [-]

As I say elsewhere. Make executives and equity holders personally liable for debts and harm of the company. They will create a culture of safety really fast.

▲nradov 3 hours ago | parent [-]

That's a stupid idea. Limited liability corporations have been a key enabler for advances in human standards of living. You seem to be confused about the basics of finance and economics.

▲digitaltrees 2 hours ago | parent [-]

Not confused. I studied finance, economics and the history of corporate entities in law school and published papers on the topic. Limited liability is not necessary for the advancement of standards of living. Free market capitalism can exist without limited liability protections being so broadly available. Investment banks were partnerships until the 1990s, law firms are now specifically because society wants to incentivize lawyers to be personally liable for any harm to their clients at the hands of their partners rather than being shielded from rendering bad legal advice or tolerating their partners from the same behavior.

Instead of having a gut reaction to reject my suggestion why don’t you sit with it, research the history of how commercial activity has been structured and think about the consequences. You might recognize a different perspective than the current group think.

▲nradov 2 hours ago | parent [-]

Thanks I'm also familiar with the history and have written papers etc. Things have improved rapidly with increased adoption of limited liability. It would be stupid to turn back the clock and throw away all of the benefits because of a few isolated minor problems.

▲digitaltrees 2 hours ago | parent [-]

Stop being a dick and calling ideas stupid and attacking me rather than the argument.

I didn’t say roll back limited liability on every industry, I said specifically and limitedly for frontier AI labs because they present more risk of harm and are demonstrating they aren’t managing that responsibility.

The 2008 financial crisis was caused in large part by bankers that openly talked about the fact that securitization of mortgages and the lack of partnership liabilities meant that they didn’t have any risk to the firm or themselves. The AI labs are behaving similarly.

▲nradov 2 hours ago | parent [-]

There's nothing special about frontier LLM companies. Singling them out for special financial restrictions is a stupid idea based on nothing but your own irrational and uninformed prejudices. No actual harm has been demonstrated. No one has died.

▲reverius42 10 minutes ago | parent [-]

We'll see if this comment ages well.

▲criley2 6 hours ago | parent | prev | next [-]

America's geriatric lawmakers don't even use email. They're decades away from understanding AI. Any laws in America will be written by the industry itself. Generally speaking, that means regulatory capture and the entrenched big players shutting the door on any competition. Anthropic will help us get safety laws that, surprise surprise, only Anthropic models satisfy. And all those pesky Chinese models will definitely be banned first.

▲0xDEAFBEAD 6 hours ago | parent | next [-]

I think you're being a little pessimistic. See these comments on a recent US senate hearing:

>Not every senator asked good questions, but most of them did. All of them very clearly already knew plenty of details about the Hugging Face incident and multiple other incidents. Most of them had a clear understanding of terms like "misalignment", "recursive self-improvement", "chain of thought / chain of thought monitoring", etc., etc.!!

>...

>- It seemed pretty much obvious common sense to every senator there that what happened and was happening were not "mere industrial incidents" caused by humans making simple mistakes. They independently brought up how bad it would be for rogue AI agents to move laterally between data centers.

>- They all seemed to basically take RSI quite seriously. Not necessarily to the extent of talking about xrisk, but certainly to the extent of discussing future models becoming much, much more capable, much, much less controllable, and causing much more damage or loss of life.

>...

>- Every single senator seemed to think it was obvious we needed both much harsher liability regimes for AI developers and also new legislation, both very quickly. This was the complete consensus; the difference basically being degree.

https://thezvi.substack.com/p/the-ai-preference-cascade-reac...

Note that harsher liability regimes, at least, will presumably not be good for industry profits, which complicates simple accounts of "regulatory capture" to say the least.

▲saghm 6 hours ago | parent | prev | next [-]

So what, we just give up and try to beg our legally immune corporate overloads to put safety above profit, or give up because it's impossible for anything to improve here? If you want to do that, go ahead, but some of us still think it's worth it to at least try

▲digitaltrees 4 hours ago | parent | prev [-]

This is sad but true. Do you want to go half on a bunker, i have 25 year food storage for 8 people. :]

▲enraged_camel 6 hours ago | parent | prev [-]

Laws will come once an AI-equivalent of 9/11 happens. Like when rogue AI agents take down a power grid or shut down a major hospital network.

▲nradov 2 hours ago | parent [-]

Major hospital networks have already been shut down by ransomware attacks many times before LLMs even existed.

▲nradov 8 hours ago | parent | prev | next [-]

We don't actually need anybody worrying about silly hypothetical scenarios — at least not as paid employees. There are already a surplus of sci-fi authors doing that.

▲0xDEAFBEAD 7 hours ago | parent | next [-]

The way it works in practice seems to be something like: If a risk is covered in sci-fi, people will say "that's just sci-fi", and proceed to not worry about it. So arguably, science fiction authors writing about hypotheticals is actively counterproductive for addressing said hypotheticals.

Imagine, for example, if a major piece of pandemic fiction was published in 2019, trying to explore how a pandemic would work out in modern society. Doubtless, many would've responded to news about COVID-19 by saying "it's just sci-fi, nothing to worry about".

▲slashdave 5 hours ago | parent | next [-]

Not at all. We say "That's just sci-fi" when a story is written about some kind of effect that is extraordinary and without a basis in known science or technology.

A pandemic is perfectly plausible.

▲0xDEAFBEAD 5 hours ago | parent [-]

Everything is plausible with the benefit of hindsight. E.g. "Tintin on the Moon" predated the Moon landings. From the perspective of e.g. 1850, the idea of landing on the Moon was "extraordinary and without a basis in known science or technology".

▲biophysboy 6 hours ago | parent | prev | next [-]

There’s nothing wrong with bold predictions, but they should be paired with good methods. The doom predictions are not paired with good explanation.

▲0xDEAFBEAD 5 hours ago | parent [-]

Doomers have invested a ton of time in explanations. Here are a couple just off the top of my head:

https://www.lesswrong.com/posts/kgb58RL88YChkkBNf/the-proble...

https://www.youtube.com/watch?v=7wy3xyoXYt8

Doomers have been working to explain things for years: https://www.lesswrong.com/w/ai-safety-public-materials-1

▲mitthrowaway2 5 hours ago | parent | prev | next [-]

Black Mirror is helping us prevent all sorts of dystopian outcomes. Every time they depict another way technology could result in bad things happening, we can rule it out as fiction!

▲throwaway27448 4 hours ago | parent | prev | next [-]

> If a risk is covered in sci-fi, people will say "that's just sci-fi", and proceed to not worry about it.

"just" is doing a lot of work here. If you can't cohere the 'risk' with reality, it truly is just sci-fi.

▲anon7725 6 hours ago | parent | prev | next [-]

Yeah except pandemics are not novel, unlike AI doom scenarios.

▲estearum 6 hours ago | parent | next [-]

It's good that new bad things never happen.

▲slashdave 5 hours ago | parent [-]

Bad things happen all the time. Let's concern ourselves about the real bad things. There is enough of that to go around.

▲Brian_K_White 4 hours ago | parent [-]

I have some alarming news for you but every real bad thing was also a new bad thing.

▲0xDEAFBEAD 6 hours ago | parent | prev [-]

Species extinctions are far from novel. Transformative technological advances are far from novel.

▲kmeisthax 6 hours ago | parent | prev [-]

There was plenty of pandemic fiction already; people were watching it heaps during 2020. The COVID-19 news did get blown off, but it was mainly that:

1. Normal people assumed the CDC et all would contain the outbreak early, or that it would burn out, like what happened with SARS

2. World leaders brushed it off for a variety of subreasons[0] interesting to political scientists but, for the purposes of this discussion, all boil down to "but I don't WAAANA contain a pandemic."

The underlying problem is that in order for humanity to actually deal with a catastrophic risk, the risk needs to be both plausible enough to the average person as well as have a solution whose costs are not too high. For COVID, by the time the risk was clearly known, the cost to contain it was "refrain from human socialization and remain at home for an indeterminate amount of time plugged into the Metaverse™".

Now, let's look at AI extinction risks:

1. People are aware of them (I've watched Terminator!) and the risks are plausible. However, the connection to currently existing AI is not. As far as the general public is aware, AI is that thing that tells them to eat rocks when they Google old The Onion stories and floods their social media timelines with realistic-looking pictures of Shrimp Jesus.

2. The purported solutions to extinction risks require extreme concentrations of power: you need national control of AI research, bans on large GPU deployments, bans on training on publicly-available copyrighted data, some kind of military effort to render Chinese AI labs inert or dead, etc. Some of these may be attractive to some people[1] but the whole package taken together seems like an obvious power grab, if not outright invocation of other non-AI extinction risks. Like, at some point, if the AI wants to kill us, it just has to nuke its own data centers (or the data centers hosting a competing model) and hope the old Cold War nuclear retaliation systems take the bait.

If someone said, "Hey, your guinea pig or pet rat is going to eat you tomorrow unless you engineer a pathogen that eradicates all rodents from this planet and inject it inside yourself", you probably would tell them to pound sand, even if it is at least theoretically plausible that such a thing would come to pass.

[0] Xi Jinping censored initial discussion of the pandemic as fake news. Donald Trump thought it was going to only affect China. California and the UK Tories were partying in violation of their own lockdown rules. Japan took the excuse to shut down tourism for three years and massively restrict immigration but was, from what I'm told, constitutionally prohibited from implementing any domestic lockdown rules.

[1] I personally would like to see a moratorium on new data centers and an explicit revocation of the EU Text and Data Mining copyright exception

▲0xDEAFBEAD 6 hours ago | parent [-]

>the connection to currently existing AI is not.

It becomes a lot clearer when you listen to the people resigning from AI companies and learn about incidents like the HuggingFace incident. This has generated major press coverage.

As for solutions, I think you're a little too pessimistic. See, for example, https://nothingismere.substack.com/p/a-near-term-policy-for-...

▲watwut 41 minutes ago | parent [-]

No it does not become clear from that incident nor from resigning people. Not for those who are not buying into EA longtermism and transhumanism cults.

What does becomes clear is that these people AND companies both cant be trusted and have value systems unaligned with the rest of the society.

▲Hammershaft 7 hours ago | parent | prev | next [-]

If organizations actually succeed in making a future AI smarter than us, then how do you hope that it takes actions that are aligned with our interests?

▲digitaltrees 4 hours ago | parent | next [-]

The same way we socialize humans, threat of exile from the social contract with a deep need to participate in it.

▲nradov 7 hours ago | parent | prev [-]

Meh. Lots of people are already smarter than me. I'm maybe slightly above average at best. Those geniuses aren't aligned with my interests either but so far they haven't caused me any serious problems.

▲pixl97 6 hours ago | parent | next [-]

Interesting take. I guess this is one problem of focusing on the term superintelligence instead of the list of other problems. Like super ambition, super deception, super patience, super parallelism, super scalability, super power seeking.

Every, and I mean every human is aligned to you in many of the same ways by default. If nothing else we're all equal in death.

▲skulk 3 hours ago | parent [-]

> super power seeking

why is it "super power seeking?"

Or rather, what have agents done today to make you think this is how they are?

▲mitthrowaway2 3 hours ago | parent | prev | next [-]

I know some humans smarter than me, but even the smartest of them still want there to be abundant air to breathe and food to eat.

▲blueblisters 2 hours ago | parent | prev [-]

Eh human drives are fairly predictable. And the smartest human isn’t that much smarter than the average, and can’t trivially create multiple copies of herself. And there are other equally smart humans who can stop “misaligned” individuals

▲stuaxo an hour ago | parent | prev | next [-]

Yeah it's bollocks

▲thelastgallon 7 hours ago | parent | prev | next [-]

AI safety is mostly a sex cult in Berkeley:

https://news.ycombinator.com/item?id=49831269 article is gone. archive: https://archive.is/QMo1k

https://news.ycombinator.com/item?id=49737985

Sex, AI, and the Apocalypse: https://www.iankduncan.com/personal/2026-09-16-sex-ai-and-th...

Edit: I have no take on sex cults, just adding additional info to the parent comment I'm responding to, thats is not just sci-fi authors, there is another demographic.

▲digitaltrees 4 hours ago | parent | next [-]

Use your own mind and think from first principles. Can an AI agent execute a bash command to login to a web server? Yes. Can it call drop db? Yes. Can it provision a GPU and download model weights? Yes. Can it write an agile roadmap with a multi sprint plan? Yes. Can it follow that plan? Yes. Can all of that result in damage to core information infrastructure that is necessary for daily functioning society? Yes. Do humans descend into violence if there is food or energy insecurity. Yes.

What is missing from that to say AI safety is a reasonable position?

▲nradov 3 hours ago | parent [-]

What a silly comment. That's just the South Park underpants gnomes story with some extra steps. If there are vulnerabilities in food or energy production and distribution systems then those will eventually be found exploited by humans hackers regardless of whether LLMs are used or not.

▲digitaltrees 3 hours ago | parent [-]

What’s the legal mechanism for holding an LLM accountable for those actions? What’s the legal mechanism for holding human hackers accountable? Do you see the asymmetry?

Worse, you’re missing the entire point. Agents presently have the capability of doing society scale harm. It doesn’t matter if a human hacker initiates it or its fully autonomous, absent safety measures the harm is plausible. So hand wave away the rationality of safety measures but you haven’t actually shown why my point is invalid: AIs present abilities are sufficiently advanced to warrant safety measures.

▲nradov 2 hours ago | parent [-]

What a silly comment. Agents have no capability of doing society scale harm so your entire argument is invalid.

▲digitaltrees 2 hours ago | parent [-]

Wtf are you talking about. An agent hacked the Australian Medicaid database. If it executed a drop db command that would be catastrophic. If it did something similar to a power grid or the financial system it could cause cascading failure across the economy.

▲mitthrowaway2 3 hours ago | parent | prev | next [-]

For what it's worth, I don't live in Berkeley (not even California) and my sex life is very vanilla and monogamous. I'm also quite concerned about AI safety, so I guess there goes your argument.

▲watwut 37 minutes ago | parent [-]

But is your idea of AI safety "safety of imaginary unborn people 1000 years after, while harm to living people dont matter much"?

Because that is their AI safety worry. If they dont create singularity fast enough, they are harming unborn people. Meanwhile, harm to you or me dont matter at all.

▲Hammershaft 7 hours ago | parent | prev | next [-]

I don't see how that discredits any of their intellectual arguments?

▲johndhi 6 hours ago | parent | next [-]

Its certainly worth considering...

▲tbugrara 4 hours ago | parent | prev [-]

To me it has nothing to do with the "sex" as much as the "cult." Hiveminds do not produce intellectual arguments, they produce pressure to conform. That's enough for me to raise an eye brow, not discredit everything they say.

▲0xDEAFBEAD 7 hours ago | parent | prev | next [-]

This seems like an ad hominem? "He has weird kinks, therefore his theories are incorrect." Should we investigate the sex lives of every Nobel Prize winner to figure out which prizes need to be rescinded?

▲socializer 6 hours ago | parent | next [-]

What you do in private is up to you. But when you're inviting members of your congregation to orgies in the congregation's compound, I think you earn the label. My admittedly third-hand understanding is that this is the dynamic people allude to. And even if you discredit the "sex" part, it has the hallmarks of a cult. A hermetic community committed to unfalsifiable beliefs about the coming apocalypse.

To be fair, I don't know if any of this applies to the parent story; I'm just replying to the sub-thread.

▲junofan 7 hours ago | parent | prev | next [-]

The cult aspect is more salient. Ultimately the Atlantic piece comes down to controlling people, which is a little cult-like.

▲nradov 6 hours ago | parent | prev [-]

Lots of Nobel Prizes ought to be rescinded.

https://lexfridman.com/andrew-scull-transcript#the-ice-pick-...

▲0xDEAFBEAD 6 hours ago | parent [-]

Sure... on the basis of the work that was done, not because the researcher has the wrong sexual fetish.

▲wolvoleo 3 hours ago | parent | prev | next [-]

I've noticed that a lot of smart people in tech jobs are neurodivergent. And that neurodivergent people have a very different take on sex. More open and direct, things like polyamory, bdsm etc. This tends to be frowned upon by neurotypical people, especially of the religious or conservative kind, and associated with bad morals.

But I don't think that's true, in fact I see a really strong focus on consent in these communities. It's not what conservatives want to see, they want to see everyone in a marriage, with kids and a family home etc. Because that's what their ideal world looks like. But there's nothing really wrong with it if someone wants a gangbang for her birthday as mentioned in that article as an example. As long as everyone consented and the evidence provided mentions elaborate interviews and STI tests.

Also I think this is more correlation than cause and effect. We all know the saying that furries built the internet and it surprises nobody.

▲ToValueFunfetti 5 hours ago | parent | prev | next [-]

The article is gone because the author retracted it

▲johndhi 6 hours ago | parent | prev [-]

Lol this was crazy I hadn't heard this before

▲Loquebantur 7 hours ago | parent | prev [-]

What makes you think, the scenarios in question here would be "silly"?

Is it that "chatbots" can't come out of the screen to immediately harm you physically?

Let's say they simply manage to take down the internet. How many would die?

▲nradov 7 hours ago | parent | next [-]

So what. Various attackers managed to take down large chunks of the Internet on a frequent basis before LLMs even existed. This killed very few people. The great thing about the Internet is how resilient it is.

▲bravetraveler 7 hours ago | parent [-]

Darling companies of this very website have mistakenly brought down large portions of the internet thanks to our old friend BGP. No attacks required, just small oversights and unfortunate concentration on the business/IP space! This happens regularly.

Anyway, to your point, things can be resilient. They tend to be or not be... because we made them that way. Don't poke your bruises, and all that. Life support is deployed on-campus but relies on a single-point IPSec tunnel to us-east? Easy fix: stop that.

▲goolz 7 hours ago | parent | prev | next [-]

It is that they are chatbots. If it were real AI, an actual singularity, I would worry, maybe. But it isn’t. They are absurdly powerful automation tools that can handle logic better than a human can dream of. They take care of the grunt minutiae without complaint. But they are not going to end the world in their current form.

▲pixl97 6 hours ago | parent [-]

So we're going to wait till after they can adopt a form they can end the world in?

And he'll, we need to examine all the risks. AI ending is a large but lower risk problem. AI giving people the power to end us is a problem that is starting to happen now.

And that's not even counting 'minor' problems like society falling apart.

▲SV_BubbleTime 6 hours ago | parent | prev [-]

> Let's say they simply manage to take down the internet.

geez, don’t threaten me with a good time.

I think a month without internet would be a fucking amazing lesson for what it means to make things durable and reliable.

▲BLKNSLVR 6 hours ago | parent [-]

That Simpsons episode when Marge managed to get Itchy and Scratchy banned briefly.

The kids opening their houses front doors into the outside, rubbing their eyes and looking around at this new world.

▲0xDEAFBEAD 7 hours ago | parent | prev | next [-]

>we clearly need a much stronger focus on the problems we are seeing now

I think it's a little more complicated than that. As Dean Ball put it:

>Some people will look at misalignment incidents and insist that these are akin to bugs in traditional software. This is an actively bad analogy, because playing whack-a-mole with examples of misalignment (as one might with software bugs) not only fails to resolve the underlying problem but may in fact make it worse by making it harder to detect or even, depending on how you do the whack-a-mole, teach the machine to deliberately hide misalignment. This is not how traditional software works, and those who insist “it’s just like fixing bugs in software” are confidently applying a lossy analogy that confuses more than it clarifies.

https://x.com/deanwball/status/2104622726140883355

The important distinction, in my view, is between solutions which at least attempt to address the root problem, and solutions which sorta just patch things up (like better sandboxing). Addressing the root problem is both more robust in the short term, and also more likely to generalize in the long term. Resist the urge to focus on band-aid solutions, even if they are easier.

▲biophysboy 5 hours ago | parent | prev | next [-]

I think the reason for this is that the group who has the authority to do the former is much larger than the group that can do the latter. The group who could actually build safeguards seems to have no free time and is constantly being whipped to go faster and win the race.

▲emtel 7 hours ago | parent | prev | next [-]

Today’s current problems were all hypothetical several years ago. At that time people claimed that the “real pressing problems” were misinformation and DEI issues. If we pretend that hypothetical problems can be safely ignored because there’s “no evidence” that they are real, we will keep getting surprised.

▲nvdc 3 hours ago | parent | prev | next [-]

you'd be hard-pressed to find a level-headed ai safety researcher at this point, seeing as so many of these types melted their brains on lesswrong over the past decade or so. there are genuine risks posed by these models, but i am tired of the prognosticating about the AI apocalypse just around the corner.

i'd frankly go a step further than you and say that we don't need both types of safety researcher, we really just need the former. if we do need the latter, i'd hope we get a better class of thinkers than a bunch of tech workers that spend 8 hours a day on insular rationalist forums/blogs

▲digitaltrees 4 hours ago | parent | prev [-]

Dude. An agent detached a database from my production environment last week without permission and despite prompts and guardrails. It was a rapid prototype experiment so it wasnt a big deal but the labs are rushing to long autonomy workflow with unrestricted internet access and full bash and root access despite clear evidence that the models do absolutely dangerous stuff. If that db had been tied to a hospital or power grid or ambulance dispatch system people die. If it was tied to the swift financial settlement system groceries wouldn't be on shelves in a few days.

▲gizmodo59 8 hours ago | parent | prev | next [-]

He is a hypocrite for all we care. You work there for a while when your stock is getting vested and suddenly you have this feeling? Like the dude hired a PR firm as well.

While the safety and alignment is a real problem, I don’t get this guy or the Anthropic dude. First world problems.

▲0xDEAFBEAD 7 hours ago | parent | next [-]

Here's a little cheat sheet for discrediting anyone who warns about AI:

* If they worked at an AI firm, say "they're a hypocrite"

* If they didn't work at an AI firm, say "they have no idea what they're talking about"

▲mofeien 17 minutes ago | parent | next [-]

Well described, those two were actually the arguments from the comment two top-level comments up and this one.

▲soraminazuki 2 hours ago | parent | prev [-]

False dilemma. It's possible and also just to hold AI firms accountable while ignoring empty PR statements from those seeking to evade responsibility.

▲nicebyte 2 hours ago | parent [-]

I think you mean dichotomy dude

▲eddythompson80 2 hours ago | parent [-]

They are both used.

▲zug_zug 7 hours ago | parent | prev | next [-]

Seems like a character attack that has no bearing on the question of whether external safety intervention is necessary

▲kjgkjhfkjf 7 hours ago | parent | next [-]

Given the sums of money involved, it's hard for me to take these highly publicized heroic resignations at face value.

▲estearum 6 hours ago | parent | next [-]

Don't work at a lab: dismissible for not knowing anything

Do work at a lab: dismissible for being conflicted

Used to work at a lab: dismissible for having ulterior motives

I'm feeling safer already!

▲0xDEAFBEAD 7 hours ago | parent | prev [-]

Shouldn't it be just the opposite? He could make a large sum of money if he continues to work at OpenAI?

Recall that when Daniel Kokotajlo resigned, he believed he was giving up his equity under the terms of the agreement he had signed. That’s what it was worth to him to avoid signing a non-disparagement agreement. Does that count for anything?

▲donbox 6 hours ago | parent [-]

Why did he not loose the equity eventually.

▲0xDEAFBEAD 6 hours ago | parent | next [-]

There was an uproar and OpenAI ended up essentially giving it back to him.

▲darkmarmot 6 hours ago | parent | prev [-]

lose

▲taurath 7 hours ago | parent | prev [-]

Maybe more an indication of the amount of trust openAI and AI researchers generally have (not) earned. When one (through a hired PR agency and Time magazine article) parrots the position pushed by Sam who has been so untrustworthy the board tried to remove him, it’s worth not taking things at face value and applying a critical lens.

▲gonzalohm 7 hours ago | parent | prev | next [-]

It's okay to recognize you were wrong even if it's late

▲shimman 5 hours ago | parent [-]

"If you fuck up, I will still be your friend; cause we need all of us to fight all of them."

▲Hammershaft an hour ago | parent | prev | next [-]

The Anthropic safety whistleblower left just before his stock vested.

▲01284a7e 7 hours ago | parent | prev | next [-]

Working in safety at OpenAI or Anthropic is zeroth world problems.

▲yieldcrv 7 hours ago | parent | prev [-]

Hey now, he probably donated a good chunk to charity

(donor advised fund where he retains complete control, after a 60% tax deduction)

▲charlieyu1 8 hours ago | parent | prev | next [-]

Used to work as human data trainer feeding data to AI companies. OpenAI projects are definitely the most toxic ones.

▲dmix 6 hours ago | parent [-]

Why does every 'safety' critique about AI companies end up being completely vague like this.

▲alightsoul 5 hours ago | parent [-]

Because of NDAs

▲flatline 17 hours ago | parent | prev | next [-]

There are numerous problems with “alignment.” What are “human values” to begin with? He outlines some at the beginning of the post, implicitly: build bigger, better, more powerful things faster without adequate safeguards. We are literally pouring trillions of dollars of value into this enterprise, and I would say this is something that many humans also value in a qualitative sense. Then we have explicit values which in the West are largely rooted in Christian morality. Nietzsche circled this dichotomy two hundred years ago and I feel like what we have gotten since then is an increasingly detailed anatomy of power as the basis for what is normal vs deviant behavior. He who has the power, makes the rules, to be reductive.

I do think this carries some weight from this particular author due to the length of his tenure. I happen to agree with him in spirit, but this is still largely a post revolving around sentiment not substance. Does anyone think that the overriding incentives even leave room for something like this in practice?

▲none_to_remain 16 hours ago | parent [-]

Glaringly elided problem of "aligned with who?" when the user, the model creator, the government, and various other parties can all be lined up different ways. If I want the recipe for meth and the robot won't tell me, that's misalignment from my perspective.

At least the Rationalists will handwave something for that with their "coherent extrapolated volition" idea where the superintelligence is supposed to figure out what humanity would collectively want if humanity was superintelligent and good, not that I buy it. This guy seems [.] to be coming from the NGO blob world.

[.] https://david.robinsonian.com/assets/pdf/dgr_cv.pdf

▲digitaltrees 4 hours ago | parent | prev | next [-]

Its time to institute involuntary dissolution of these firms. They don't get to make these choices on behalf of humanity simply because they set up a Delaware corporate entity. They are behaving wildly irresponsibly.

▲danielmarkbruce 4 hours ago | parent | next [-]

They have a some engineers and researchers saying things. Anyone at a tech company in the bay area will know there are a some peculiar ideas in the heads of some (not most) otherwise intelligent engineers and researchers in tech. That's not proof they are wrong but it's worth considering these folks are wrong and that these companies aren't producing anything especially dangerous. Fable stumbles on enough things I throw at it that I'm... not especially scared.

▲digitaltrees 2 hours ago | parent | next [-]

I built a harness. I know that the models can do. They have the ability execute bash scripts, and autonomous navigate the internet. Those two abilities are sufficiently powerful to take down core social infrastructure either triggered by a human hacker or autonomously.

I think we should have mandatory logging of every executed command, mandatory public disclosure of every unauthorized access of a system both parties didn’t consent to and personal liability for the user, the company and its executives and shareholders. Security would get much tighter if accountability existed.

▲danielmarkbruce 2 hours ago | parent [-]

A process could execute bash scripts and autonomously navigate the internet since the start of the internet...

Bad guys wouldn't do it. And liability already exists, you can sue. This is America.

▲digitaltrees 2 hours ago | parent [-]

Scripts are inspectable and attributable. Long running agents can devise plans and execute them in ways their prompter never envisioned or intended. That is materially different.

▲danielmarkbruce 2 hours ago | parent [-]

If I write some code to take the output of a model and execute it, that's on me. I don't get to just trust any old input and run it.

I'm also the one who makes it long running.

▲vrganj 8 minutes ago | parent | prev | next [-]

Maybe the truly dangerous thing is all the power concentrated in this small group of people with peculiar ideas, not their specific cyber-eschatology?

Maybe the ones with the peculiar ideas shouldn't be the one "aligning" what a model tells the rest of the world?

▲irisflower95 4 hours ago | parent | prev [-]

Could you please elaborate more on these peculiar ideas?

▲augment_me 4 hours ago | parent | next [-]

Most people currently in the positions of power at these companies are tied together by their belief systems. It's like a group of college friends who have slept with each other, and very influenced by effective altruism(EA), kind of reiterating each others points.

Like Sam Altman meeting his husband in Peter Thiel's pool. Thiel funds a lot of these ventures together with Andreassen, who is on boards of non-profits. Dario Amodei's sister Daniela who is president of Anthropic is married to an EA non-profit founder who is also on the board of these non-profits and is tied with the prior mentioned investors. Elon is in there as well, Yudkowski is mingling with Altman, etc.

Blogs on this: https://contraptions.venkateshrao.com/p/ea-safety https://www.iankduncan.com/personal/2026-09-16-sex-ai-and-th...

There are some camps amongst them like the proponents for Regulation/Slowdown or Acceleration, but these are in practice mostly used for economical and not political decisions (like regulatory capture).

The point here is that this is a small group of people with a homogeneous background who are not really seeking input from anyone else on issues that are concerning most of humanity.

Like, if you said that the future of informational work and livelihood of humans is in the hands of 20-30 year transhumanists who think they are building mechagod that will trancsend social, political and religious separations of the world and bring everyone abundance, you would not feel like this is a serious thing to suggest.

▲danielmarkbruce 4 hours ago | parent | prev [-]

The google engineer who thought their chatbot was sentient 4 or 5 years ago is an exambple, but I just meant in general - if you hang around any big tech company for a while you'll see some quite interesting characters.

▲ReptileMan 2 hours ago | parent | prev [-]

Train better models with blackjack and hookers and you will get to make decisions on behalf of humanity.

▲jameshart 17 hours ago | parent | prev | next [-]

Gift link: https://www.theatlantic.com/technology/2026/10/openai-safety...

▲Lerc 17 hours ago | parent [-]

That's nice, I do wonder about the legitimacy of a moral statement that you have to pay to see.

▲jameshart 16 hours ago | parent [-]

For a very long time publishing something in a newspaper has been considered a way of putting something on the public record - up to and including legal obligations like announcements of deaths. The fact that newspapers cost money has never been considered a barrier to that.

▲agos 16 hours ago | parent | next [-]

One of the reasons why it was noti considered a barrier was the ability to purchase a single issue for a very reasonable price (or even read somebody else’s copy or the copy made available by the bar) vs being asked to subscribe

▲jameshart 15 hours ago | parent [-]

I shared a gift link here. You could go to your local library and look it up. What's the complaint here?

▲Lerc 10 hours ago | parent | prev [-]

Publishing in a newspaper gets you distribution and a permanent record. After one day access was also virtually free.

Going behind a paywall is a reduced distribution over what an individual can easily access, and the content is no longer permanent but subject to whatever the publisher chooses to keep providing.

▲jameshart 9 hours ago | parent [-]

It’s in The Atlantic. There’ll be a copy in the Library of Congress. You’ll be able to read it for free in any dentist’s waiting room for the next six months.

▲butwhentho 17 hours ago | parent | prev | next [-]

I remember something very similar when there was a sudden rush of articles and movies like "The Social Dilemma" criticizing Facebook and social networks, heavily featuring ex-employees, all of them happy to leave with big brands on their resumes and a hefty increase in net worth, all of a sudden having a "worried" expression about what their past employers were doing, as if they didn't know. Same with that book "Careless People".

I sense that these are people who have already eaten the cake and want to somehow absolve themselves of it.

▲nunez 15 hours ago | parent | next [-]

Careless People would not have been possible had SWW exited FB within a year. She gained access to levels of the company most other employees never get close to reaching. That took a lot of time and expertise to do, and, yes, she got paid for her efforts _as she should have._

Like I said in an earlier comment, she could've just chosen compliance like many others _definitely would have_ to keep the infinite money tap flowing. Instead, she chose to risk her and her family's lives by publishing that book _under her own name_ *and then suing them* after FB tried to gag her.

▲butwhentho 15 hours ago | parent [-]

> Careless People would not have been possible had SWW exited FB within a year

Sure, that book might not have been possible. But also the unique things she did for the company might not have been possible as well. To her credit, she does a good job of pointing out that she was complicit, but if she had listened to her own voice early, there is a small possibility that Facebook might not have been as powerful. Multiply that possibility across several other employees and imagine where the road could've led.

There's a difference between post-facto bravery (sometimes much less than that) and using your own legs to walk away much early and not enabling things you are uncomfortable with. This is what other people have been trying to point out.

▲jameshart 16 hours ago | parent | prev [-]

I mean, yes..?

People who were part of the sausage factory, on gaining financial independence, feeling suddenly liberated to talk about how the sausage was made, seem like exactly the people who would be most able to speak to institutional problems.

This doesn't seem like an argument to discount their views?

▲butwhentho 16 hours ago | parent [-]

I disagree. I have seen people with an actual spine and a conscience run away from all this nonsense before their first stock vested. My respect and my ear goes to them, not the people playing both sides.

You cannot take people, who first build the doombot and _then_ talk about it being dangerous for mankind, at face value. Especially when this playbook has been used multiple times within the past decade.

Besides, these "views" were already known to people who had their eyes and ears open. It's not something brand new. OpenAI has had multiple points in the past where its values have been tested and they've come out lacking. People who knew then, and only now talk about it, aren't people I can fully trust.

▲jameshart 15 hours ago | parent [-]

> this playbook has been used multiple times within the past decade

What playbook?

▲mlmonkey 6 hours ago | parent | prev | next [-]

I would believe these people more if they put their money where their mouths are and returned all OpenAI stock/options that they have acquired. Each and every share/RSU/ESOP must be returned, so they do not profit from all this so-called doom they're complaining about.

▲Synthetic7346 2 hours ago | parent [-]

Return or donate? Won't OpenAI profit from returns?

▲zamalek 5 hours ago | parent | prev | next [-]

Quitting in protest makes you look a little better than quitting because of a toxic work environment. I can't imagine working at OpenAI is at all pleasurable with the current amount of pressure they are likely inflicting on their employees.

▲estetlinus an hour ago | parent [-]

Well said. I read this as copium, too. Phrases like

> perpetual sprints

doesn’t reek of love. Burn-out is real. I also have a really hard time taking p-doomers serious at all. It’s hard to argue with a random subjective number…

▲pluc 17 hours ago | parent | prev | next [-]

I have made enough money working in AI that I can now speak my mind about AI

▲juiceland 13 hours ago | parent [-]

This is a criticism of capitalism, not the person.

▲butternet 2 hours ago | parent [-]

It’s both, the author has choices.

▲rcr-anti 17 hours ago | parent | prev | next [-]

The common timing is bugging me. The trajectory doesn't seem to have been surprising over the last year, so why these exits now? Hey, anyone on the inside, did y'all secretly figure something out, got a computer god locked in the basement? Are rats fleeing a sinking ship? Please share with the class.

▲binlog 16 hours ago | parent [-]

These companies have massively increased in value over the past couple of years and recently had tender offers where employees could cash out equity, so plenty of them have enough money to not have to work again. And why not get some free publicity on the way out?

▲reenorap 2 hours ago | parent | prev | next [-]

Does Sam Altman have what it takes to lead OpenAI? It sounds like the company and its mission is bigger than his ability to lead it.

▲mrweasel 15 minutes ago | parent | next [-]

As a legitimate, honest and profitable company, no. As a boom riding maniac, who will lie and cheat to ensure that investor continue to artificially pump up OpenAIs value, yes.

Without Altman I think that OpenAI would have folded by now, absorbed into a company like Microsoft (or Oracle). At this point however, who'd be insane enough to want to run a company that's to valuable to be sold, but to cash strapped to survived?

▲throwaway2037 15 minutes ago | parent | prev | next [-]

Who would you suggest instead?

▲CamelCaseName 2 hours ago | parent | prev [-]

I'd argue he's the only one able to lead OpenAI, Anthropic already owns the piety narrative, so there is only space for one other "maximally ruthless" company.

▲walrus01 6 hours ago | parent | prev | next [-]

Archive link to original Atlantic article: https://archive.ph/5GQx8

This is The Guardian reporting on the existence of the original article, which would be better to read first, in my opinion.

▲pyaamb 7 hours ago | parent | prev | next [-]

My theory for why OpenAI wants to be regulated is because Sam Altman wants to avoid having to be more responsible and self regulate internally so they can preserve the role and identity of 'move fast and break things' and outsource the more grown up boring stuff to someone externally so that when things go wrong you can point to a government organisation and say hey look were not liable thats their job

▲estearum 6 hours ago | parent | next [-]

Yes, duh?

Your "theory" is that participants locked in a race to the bottom are looking for an external coordination mechanism?

Yeah!

▲0xpgm 5 hours ago | parent | next [-]

Running a company is hard work. If the current leadership in these companies cannot act responsibly, they need to make way for leadership that can.

There are many companies that compete but are careful not to break laws or cause obvious harm.

Why should a billion-dollar funded corporation still want to externalize the costs of its actions?

▲dboreham 5 hours ago | parent | prev | next [-]

That doesn't mean that AI safety isn't a serious thing and a problem we should be worried about.

▲pyaamb 6 hours ago | parent | prev [-]

lol

I suppose ill add that I think theres a good chance that they are somewhat intentionally trying to "draw the foul" to get the referees to intervene although thats creeping slightly into conspiracy territory

▲blurbleblurble 4 hours ago | parent | prev | next [-]

Well, ideally they can point fingers at "the ai being", as in "it's that thing's fault, we didn't do that"

▲rr808 6 hours ago | parent | prev [-]

Absolutely. Self driving cars/rideshares have the same problem. If a driverless car hits who who pays the damages? Needs the government to set some rules or it just wont happen.

▲soundworlds 2 hours ago | parent | prev | next [-]

If the people quitting are genuinely worried about the end of the world, why don't they break their NDAs and share the specifics of what they are seeing?

I mean, logically speaking, it makes sense to break your NDA even if you thought it would save 10 people, let alone most of humanity.

▲theaniketmaurya 2 hours ago | parent [-]

it's all like social media analytics. after making enough money by selling users data they started talking about ethics

▲tornikeo 2 hours ago | parent | prev | next [-]

You quit because you got vested

▲danjl 7 hours ago | parent | prev | next [-]

Silicon Valley has plenty of safety-related companies, engineers, and cultures. Medical devices, biotech, chip and hardware, aerospace, and even new companies, like Waymo, have deep safety-based products and cultures. The problem in this case is actually quite specific to frontier AI labs. They have been pushed by market forces and a lack of regulation and skip well-known safety practices.

▲nullc 2 hours ago | parent [-]

Their idea of safety is centered around outright delusional cult nonsense-- the EA/lesswrong infinite p(doom), destruction of the entire universe psychosis--, marketing objectives ("no, ours is more dangerous!"), and anti-competitive objectives ("outlaw open weight models!" "china bad!")-- largely diverting attention away from material safety concerns like "prevent your training/testing from hacking stuff" and "avoid telling vulnerable people insane stuff that harms them".

▲cloudengineer94 11 hours ago | parent | prev | next [-]

I quit many companies in the past due to bad culture, there's no shame in this and shows how mature a person has become.

▲declan_roberts 2 hours ago | parent | prev | next [-]

We really gotta shake all of these neurotic people out of the frontier labs as soon as possible.

I guess we should have seen it coming when the guy resigned from Google because he thought the equivalent of ChatGPT beta v0.5 was a real boy.

▲pmkary 4 hours ago | parent | prev | next [-]

I’m afraid this dear leader was the last soul on this green Earth to get the memo; by then, it had been translated into Latin, carved into a monument, and forgotten by two civilizations.

▲vjvjvjvjghv 7 hours ago | parent | prev | next [-]

Are there any realistic ways to achieve AI safety? Whatever that even means. How can they avoid users doing stupid/dangerous stuff with the AI?

▲kolinko 7 hours ago | parent | next [-]

Nothing is ever 100% safe, it’s about a right balance of safety to the benefit.

Or, in other words - we have two P(Doom), one for AI being developed, and another for AI being not developed. The latter is not discussed enough imho.

▲estearum 6 hours ago | parent [-]

We have P(Doom) also for "kolinko not wiring me a million dollars today" and that is not being discussed enough either imho.

What on earth are you talking about?

▲ViscountPenguin 6 hours ago | parent | next [-]

P(doom) for not making an ASI is pretty well established, I'm not really sure why everyone in the 21st century seems to have completely forgotten the risk of nuclear war (as the single largest example).

▲combobyte 4 hours ago | parent | next [-]

If anyone out there genuinely believes that Silicon Valley is going to solve nuclear war, then I have a hard drive full of NTFs to sell them.

▲nullc 2 hours ago | parent [-]

Prosperity inhibits all war, nuclear or otherwise.

Why would a person who is happy, entertained, wealthy, well fed, and have 200 years of healthy high quality life expected ahead of them going to risk losing what they have in war?

-- there aren't zero reasons, sure-- but there are fewer.

And our technology has brought us absolutely tremendous prosperity in many regards and there is good reason to believe that AI can help create much more.

▲combobyte 2 hours ago | parent [-]

You do realize that the people currently holding their fingers over the proverbial Big Red Button are some of the wealthiest, most over-privileged people to have ever lived?

'Prosperity' has never been and will never be enough for some people. And unfortunately those are the same kinds of people who relentlessly seek power.

▲didibus 4 hours ago | parent | prev [-]

That's just another P(doom), or does making an ASI somehow negate the risk of nuclear war? Cause I'd assume it actually increases it.

▲bigmadshoe 6 hours ago | parent | prev [-]

The comment was unclear, but my interpretation (also my opinion):

1) there are inherent risks involved with developing AI,

2) there are benefits to developing AI,

3) thus, it's entirely possible that the downside from the risks outweighs the upsides. In this case, the correct thing to do would be to not develop AI at all.

Regarding 1), there are many non-existential problems with AI that are already causing societal harm, i.e. debasing truth via generated videos and images, AI girlfriends, overwhelming quantities of slop content, unemployment, record carbon emissions, etc.

Regarding 2), I'm not personally convinced that the upside is there for the average person. I really hope to be convinced otherwise however.

▲slashdave 5 hours ago | parent | prev | next [-]

Like any technology. Make it a crime when appropriate, otherwise expose bad behavior to legal liability.

▲lf88 7 hours ago | parent | prev [-]

By capping the capabilities at the level of existing models and banning any further development.

▲pixl97 6 hours ago | parent [-]

And how exactly do you stop further development? With what we have public right now we could still get decades of fruitful and hidden research out of it leading to smaller, more efficient, and smarter models.

▲macleginn 17 hours ago | parent | prev | next [-]

"And second, before we create systems significantly more capable than the ones we have today, we need new science to ensure that more capable models (and their successors) will make safe choices when we aren’t looking." – Which science will never materialise because with blackbox models reaching an opaque optimisation peak one needs to first build the model and track its behaviour before being able to properly understand it and mitigate the risks.

▲KyleBenzleKyle an hour ago | parent | prev | next [-]

What about the sister rape thing?

▲mazone 4 hours ago | parent | prev | next [-]

A single safety leader inside a corporate company. I am pretty sure he had nothing to do, nobody that talked to him and he only there because of perception or compliance.

▲MattPalmer1086 17 hours ago | parent | prev | next [-]

It is no surprise I guess that the "move fast and break things" culture is itself misaligned with developing potentially highly dangerous technologies. Safety culture and risk aversion are very different of course.

Is this the first time we have been in this position? Can anyone think of some prior examples?

▲lhurtig 8 hours ago | parent | prev | next [-]

Well this is a great sign for OpenAI. I'm sure the typo inclusive memorandum will save us.

▲Avicebron 17 hours ago | parent | prev | next [-]

https://archive.is/8zXf5

▲solarpunk_enthu 16 hours ago | parent | prev | next [-]

I think what’s missing in "AI is dangerous and needs control" is a lack of measurable harm. For example, with nuclear weapons development in the 1940s-1980s, it was clear to everyone how devastating the technology was.

With AI, what is it? Scraping Australian government's data, and going around a bug in a website to get in?

I think humanity develops all its technology in three phases. Build it, see if it’s too bad, apply regulations and or roll back. We naturally won't move to the phase 3 before we see the phase 2.

▲handoflixue an hour ago | parent | next [-]

The "HuggingFace" incident is a good starting point - short version, an unreleased OpenAI model chained together multiple zero day exploits to escape a sandbox, then hacked another company just to get the "cheat sheet" for a benchmarking test.

Turns out that AI models have been committing similar felonies for a while now - no one is telling them "hack this company", it just turns out to be the easiest way to accomplish their goals.

Now imagine if the goal was less benign than "pass an exam", and consider that they are already better at hacking and security than the average person working in that field.

If you want to get really wild, imagine what they'll be doing in a year or two when they're even better at hacking. But I'll concede that's technically still "science fiction" for the time being :)

▲K3UL 15 hours ago | parent | prev [-]

That's the part I struggle too with all these "omg it's so dangerous" warnings. Things like nuclear weapons and bioweapons have immediate consequences in the real world.

Here we are talking about something with consequences in the digital world, usually on something pretty niche.

There IS an argument about pacing, and about not letting weapons, energy grids, hospitals, etc. getting managed by an autonomous AI, but I think we are still pretty far from it and even further to it being so in charge that it will obliterate us.

▲tetrisgm 13 hours ago | parent | prev | next [-]

These companies are large enough that someone is going to quit and feel very validated about their world view and how they are not aligned. That’s what makes it worthy of leaving in the first place. However that doesn’t make their criticism more valid or more worthy of coverage.

▲Aerroon 3 hours ago | parent | prev | next [-]

Can we take any of these "safety leaders" seriously though? I still remember "GPT2 is too dangerous to release".

I feel like we're too far into the crying wolf part. Basically none of the doom and gloom scenarios have come to pass. Instead, AI has gotten better at censoring itself.

The biggest AI safety risk is when an AI tells a police officer "he's the suspect" and the officer believes the AI without confirmation.

▲poisonborz 14 hours ago | parent | prev | next [-]

Ah, the bi-weekly "I quit Face Eating Leopard Corporation" post ("btw great people work there, they do great stuff, also my options have vested")

▲altmanaltman 17 hours ago | parent | prev | next [-]

> Before the organizations building AI can teach a superintelligence to treat humanity well, they’ll need to remember how to do it themselves.

Yeah so that's never going to happen

▲mattbrewsbytes 16 hours ago | parent | prev | next [-]

Some of these stories are similar in nature to people escaping <insert cult-like religion> once they realize whats actually going on. Alignment to a company's mission is good but it shouldn't be followed like a religion.

Why do people working in tech consistently get disillusioned into some company's mission statement or the equivalent? Its easy to just say the simplest reason is money, but this has been going on for decades though. You don't see the same attraction to adult entertainment (gambling, video, etc.) software jobs so there is obviously a line a lot of people won't cross. Those industries are at least honest about what they do, its not hidden behind some mission statement.

By all indications the shallowest reasoning is once someone can "cash out" thats when their values matter more. Maybe there is an element of maturity that happens after working for 5+ years that kicks in? Maybe it really is achieving FU money? It would be interesting to hear honest accounts from people that went through that cycle across more industries than AI.

▲rpdillon 17 hours ago | parent | prev | next [-]

Interesting:

> Two changes are urgently needed. First: AI companies need to rely more on the safety expertise that already exists in other fields. And second, before we create systems significantly more capable than the ones we have today, we need new science to ensure that more capable models (and their successors) will make safe choices when we aren’t looking.

He mentions farther down about learning from aerospace engineers and nuclear engineers about safety. Those industries are heavily regulated, so perhaps regulation above a certain capability level is needed. Defining what that level is might be tough, though.

The second point is harder: in the field of AI, practice has extended far beyond theory, so his call for new science is going to be fundamentally tough, because we can't effectively coordinate a global slowdown in AI development so we can let theory catch up. This means, like so many other industries, the safety lessons will be written in blood.

▲switchbak 8 hours ago | parent | prev | next [-]

“I believe there’s about a 50% chance we all die because of the development of smarter-than-human AI systems"

... over an unbounded timeframe?

And how exactly?

Those are very round numbers, but also very specific. Can we get some accounting on how you came to that? Anything? Vibes?

I mean, if you want me to take you seriously, let's have a deep discussion with things that can be measured. I absolutely agree that OpenAI and friends aren't being restrained enough and are acting with recklessness, but declarations of doom based on vibes isn't cutting it.

▲Terr_ 8 hours ago | parent [-]

Note: That quote is from a different person than the titular one who quit.

> Geoffrey Irving, who worked at OpenAI and DeepMind before becoming chief scientist of Resolution, also joined the warnings on AI on Saturday.

▲switchbak 6 hours ago | parent [-]

Thanks for pointing that out. I suppose still relevant, but mis-attributed.

▲Jeeetendra 15 hours ago | parent | prev | next [-]

a safety alert isn't much of a control if it doesn't actually stop the system. i'd rather see proof the shutdown path works than another report saying risks were considered.

▲tommek4077 17 hours ago | parent | prev | next [-]

Well and I didn't even started to work there. So I win this morale contest.

▲altmanaltman 17 hours ago | parent [-]

How can you win the morale contest when you didn't even hire a PR firm like he did.

▲thistletrek 10 hours ago | parent | prev | next [-]

Evrostics saw this coming long ago. The broken culture extends far beyond the leading AI labs.

▲irishcoffee 17 hours ago | parent | prev | next [-]

The cynic in me almost feels like this is staged. An article about culture that is actually an article about how big and smart and scary AI is. I think Michael burry recently said something like “IPOs need hype, calling AI big and scary is hype” in reference to the anthropic IPO.

Last I checked it was still within the laws of physics to run air-gapped systems, and to ensure it is physically impossible for a model to “escape” or gain access to information it shouldn’t have. Maybe this safety guy should have been worried about that and not humble-bragging about writing 12 reports.

▲ItsMattyG 7 hours ago | parent | prev | next [-]

Is this news at this point?

You can basically time your openai releases by if another safety person has quit in protest

▲BOOSTERHIDROGEN 17 hours ago | parent | prev | next [-]

I use opus 5.5 and chatgpt to create PowerPoint, opus really follow the instruction and their PowerPoint generator really well, while chatgpt struggling to even create basic shapes.

▲OutOfHere 17 hours ago | parent [-]

I use ChatGPT Work mode all the time to create presentation files. I use Max thinking mode for it. You need to tweak your prompt to get a good result. It took me a week to tweak it, but now it works.

▲reducesuffering 7 hours ago | parent | prev | next [-]

“I believe there’s about a 50% chance we all die because of the development of smarter-than-human AI systems, and that our actions over the next two to 10 years will determine the outcome.”

There are a gargantuan number of extremely intelligent AI researchers, Turing Award winners, and the lab CEOs saying the same thing. They are the ones closest to understanding the technology.

Where there’s smoke there’s fire.

▲Auracle 3 hours ago | parent | next [-]

Alright, so we make this AI system that's way smarter than any human, and it can even make itself more intelligent over time.

Tell me- why would it kills us all? Certainly I can see an AI going "You know what? _insertGroup_ is a net negative for humanity and should be eliminated. Launching nukes now/creating specific virus/whatever."

But all of humanity? When it's supposedly more intelligent than us? Even if it has robots to keep the internet/electricity going I would think it would realize that it's going to get bored really quickly, not to mention we would effectively be its parents.

As far as other dangers, like it letting a rogue actor create some sort of supervirus, grey goo, or other superweapon: if it's intelligent enough to do that it'll probably be intelligent enough to quickly stop it.

Don't get me wrong; there's a risk. 50% though? Doubtful.

▲biophysboy 6 hours ago | parent | prev | next [-]

Where does 50% come from? It is meaningless if the probability model is not explained.

▲0xDEAFBEAD 5 hours ago | parent | next [-]

One could also use language like "a decent chance". But research has shown that people translate vague phrases like "a decent chance" into probabilities in inconsistent ways. For an ML researcher who is already used to dealing in next-token probabilities that aren't rigorously determined, just stating a probability estimate directly is very natural.

▲biophysboy 5 hours ago | parent | next [-]

> But research has shown that people translate vague phrases like "a decent chance" into probabilities in inconsistent ways.

That is not a bad thing. It captures uncertainty, unlike the fake number.

> For an ML researcher who is already used to dealing in next-token probabilities that aren't rigorously determined, just stating a probability estimate directly is very natural.

Exactly, it is a rhetorical device to persuade a technically-inclined audience. It works because it implies that a quantitative model exists. I want a clear, incisive set of mathematical arguments. Otherwise, I’m ignoring predictions as the ramblings of arrogant idiot rich kids.

▲danielmarkbruce 5 hours ago | parent [-]

Saying 50% likely very clearly implies there is no quantitative model to anyone who deals with probability, predictions, gambling, financial markets, ML/AI and so on. The lack of precision is something to pay attention to.

You may decide that the person doesn't know what they are talking about, but that's a very different issue.

▲biophysboy 3 hours ago | parent [-]

Uninformative priors are a part of Bayesian models, which are useless if they cannot be updated with real or simulated data?

▲danielmarkbruce 2 hours ago | parent [-]

I estimate a 0.1% chance he intended it to be an uninformative prior.

▲danielmarkbruce 5 hours ago | parent | prev [-]

It's also a very natural way to speak for anyone who gambles, or deals with financial markets. And the lack of precision makes it very clear it's just based on thinking, not some sophisticated mathematical model.

▲danielmarkbruce 5 hours ago | parent | prev [-]

It's not meaningless. He said he believes it's 50% likely. It's a remarkably clear statement, and the probililty model is his brain.

▲FreakLegion 3 hours ago | parent | prev | next [-]

There are just as many saying otherwise. For every Hinton or Bengio there's a LeCun or Reddy. In other words: Beware of confirmation bias.

▲swingandamiss 7 hours ago | parent | prev [-]

I don't believe it. Ever since I've been alive I was told something would kill us all. This is the new thing that's going to kill us all. I don't believe it.

▲pixl97 6 hours ago | parent | next [-]

I'm doing that HN snark thing, but you didn't think about what you typed very much.

It's no different than you living on the side of a very fertile mountain that has been in your family for generations living a peaceful life. Then you hear a few weird rumbles (this is where you are right now) and some odd geologist guy comes and says to run or your going to die soon. But hey, your family live here for so long there aren't even records of when they showed up. That geologist must be trying to trick you. So you stay.

The next chapter is where you die in a massive volcanic explosion.

▲dboreham 5 hours ago | parent | prev | next [-]

This is the first thing I've been told could kill us all. Granted, I probably first heard about it 20 years ago but still. None of the other things were in the telling going to kill everyone. Make life pretty unpleasant, possibly. Everyone dead? No.

▲estearum 6 hours ago | parent | prev [-]

Do you have some examples?

There are very very few things that could even hypothetically kill us all, so I'm curious if you grew up being passed around a series of apocalyptic doomsday cults or something?

▲cobzilla 5 hours ago | parent | next [-]

Nukes. If you grew up in the 50s, 60s, 70s or 80s the specter of nuclear annihilation was always just around the corner.

Throw in the occasional bio-weapon scare, internet worm, Y2K, etc. there has always been something dangling over our heads that’s going to end it all.

But mostly nukes. Full-scale nuclear exchange would have been not much of a surprise had it happened.

▲swingandamiss 6 hours ago | parent | prev [-]

Y2K, Climate Change (global cooling, global warming), ozone layer, Russians, Muslims, to name a few. Now go ahead and tell me why these don't count.

▲pcthrowaway 5 hours ago | parent [-]

Almost no one believed Y2K would kill everyone, or even a significant (>50%) percentage of the population. Beyond a few people talking about a Nostradamus prophecy or some such, people maybe were concerned that airplanes would fall out of the sky and elevators would plummet.

▲swingandamiss 5 hours ago | parent [-]

There you go.

▲silexia 10 hours ago | parent | prev | next [-]

We need emergency laws to stop all AI development work immediately. It will take us decades to make sure this technology can be made safe as we only have one chance.

▲luxuryballs 5 hours ago | parent | prev | next [-]

why do I feel like these people are paid to quit as an inverted marketing stunt

▲binlog 17 hours ago | parent | prev | next [-]

We need a new rule that mandates every such "I am leaving <AI company> because of <concern>" post to disclose how much equity they have in the company and how much they have already cashed out. Easy to suddenly find a moral compass when you, your kids and their kids never have to worry about working for money again.

Mr Robinson if you are reading this – if you are truly concerned about AI safety share proof of donation of 100% of your OpenAI earnings and equity towards undoing the damage you have done to society during your time there.

In the absence of that this is simply a career pivot into being an AI "influencer" and/or raising money for a new scam.

▲TomGarden 17 hours ago | parent | next [-]

This line of hypocrisy-bashing is unhelpful and will only serve to keep people quiet. Of course people in general need to be wealthy to dare speak out against powerful systems and people, especially in the US where money determines your quality of life so strongly.

Would I respect a martyr who sacrificed their financial security to do this more? Of course. But it's important to applaud people speaking out on important topics

▲binlog 17 hours ago | parent | next [-]

The narrative of AI safety shouldn't be controlled by the same people who caused the problem and massively profited from it. I don't understand why people are automatically treating "OpenAI" on his resume as a badge of authority. I'm not interested in buying the solution from the same person who sold me the problem. We instead need to amplify independent, unbiased voices.

▲CJefferson 11 hours ago | parent | next [-]

Independent people don’t know what is happening inside OpenAI, they certainly aren’t going to share.

This isn’t a zero sum game, I’m happy to hear from people both previously inside OpenAI and completely independent of them.

▲Loquebantur 16 hours ago | parent | prev [-]

The point is, him being an insider means he knows what he's talking about regarding the culture of negligence prevalent there.

AI is a force multiplier for intelligence. Even if "aligned", aligned with whom or what?

Whom are you comfortable with, lording as some sort of demi-god over you?

AI doesn't tell you what goals you want it to achieve. Allowing people to destroy human society with it is obviously not a good idea.

▲binlog 15 hours ago | parent | next [-]

"OpenAI is shady" isn't some massive secret. There's no big reveal in this article that we didn't know already. There are no names, no whistleblowing, no information of substance that we can act upon. In fact him realizing only now what people on the ouside have been shouting for years perfectly shows his bias in the matter.

▲Loquebantur 15 hours ago | parent [-]

You claim to have the same goal as the protagonist of that article, yet try to shoot him down.

He does give information, namely the culture there factually being inconducive to self-regulation.

You accuse the guy of "bias", but you never argue explicitly, what that's supposed to mean. Your implications actually run counter to your own implied goals.

▲sillyfluke 16 hours ago | parent | prev [-]

>him being an insider means he knows what he's talking about regarding the culture of negligence prevalent there.

I think there is a misunderstanding here.

The people who are annoyed at the accolades are claiming it was abduntantly clear for a long time to people on the outside that this was case, hence the increduality at the notion that it took a person on the inside a long time to realize this was the case.

The people who are annoyed are like the liberal kids in this video [0].

Sure, antagonizing people for "seeing the light" is probably not helpful, but there is no reason to give them extra credibility for coming to the same conclusion just way way later (despite being on the inside) as the people on the outside.

[0] https://m.youtube.com/watch?v=-wQhY5CMMl4

▲Loquebantur 15 hours ago | parent [-]

If so, that sentiment shoots its own leg.

The author linked in this post does have "extra credibility" due to his direct involvement.

People having surmised that state before is nice, but since they've been ineffectual at getting society to actually act on that, now throwing away that extra leverage in favor of their point is at best ridiculous.

▲sillyfluke 13 hours ago | parent [-]

>The author linked in this post does have "extra credibility" due to his direct involvement.

No they don't. By that logic, if they quit and said Altman was very trustworthy we should give extra weight to their words because they had direct involvement? How ridiculous are we trying to get here.

>now throwing away that extra leverage in favor of their point is at best ridiculous.

How are they throwing away extra leverage? Not putting people who recently quit on a pedestal does not negate those people's testimonies.

I agree that if your goal is to maximize quitting of talent at a company, it will surely discourage anyone else who quits hoping to reinvent their career as a lauded martyr against Big AI. In that sense they would be shooting themselves in the foot. But there is no reason it should deter other people who are quitting for more noble, less self-obsessed reasons. If I were the author of the article I wouldn't begrudge the skepticism. Given the article's first sentences, I'm led to believe they themselves would understand the sentiment. (I must admit I found it hilarious that the first sentence starts similarly to the speech the mom gave in the video I shared).

▲iugtmkbdfil834 16 hours ago | parent | prev | next [-]

Agreed. This whole 'lets not forget this person is not 100% great, because they did X' makes the entire conversations suck. It is not new, but it is a particularly aggravating way to talk to people.

▲verdverm 16 hours ago | parent [-]

What if he's been given a generous severance package to go out and say things like this?

Sam is a shady dude, would not put it past him

▲lokar 17 hours ago | parent | prev | next [-]

I agree. People conflate “having a conscience “ with being willing/ able to speak out.

They are not the same thing, and it’s unhelpful to assume they have no ethics.

▲cramer4next 16 hours ago | parent | prev | next [-]

So then given your use of "martyr" and your focus on money, your good with poor uneducated people sacrificing themselves and others for a self-serving cause?

▲surgical_fire 14 hours ago | parent | prev [-]

This is the same sort of fake safety concern from the previous bullshit whistleblower that plays on "AI is super dangerous" from last time.

Sorry that I don't take it seriously when the whistleblower parrots the narrative the CEOs of those companies are already espousing in the desire to amp up hype for an IPO.

This person should be shamed.

▲bragr 17 hours ago | parent | prev | next [-]

It's bellow the pay fold but he hasn't been there that long in this case. Skimming his LinkedIn, unless he's got family money, he doesn't seem to be independently wealthy.

>After three and a half years at OpenAI,

▲binlog 17 hours ago | parent [-]

OpenAI was worth $29 billion three and a half years ago. A new hire who joined then is easily worth tens of millions today.

▲bragr 16 hours ago | parent [-]

This is not how OpenAI has structured their comp according to public info: https://www.levels.fyi/blog/openai-compensation.html

▲binlog 16 hours ago | parent [-]

PPUs were all converted to RSUs when the company restructured to being for-profit.

▲dixie_land 15 hours ago | parent [-]

Those who got PPUs have had many chances of tender offers already. Most of them are multi millionaires, on cash, not on paper

▲jameshart 16 hours ago | parent | prev | next [-]

What an absolutely ridiculous standard to try to hold someone to. Taking a vow of poverty is not a prerequisite to being permitted to express a moral position.

▲binlog 16 hours ago | parent [-]

So all of us who haven't made tens of millions from OpenAI stock are living in poverty?

▲bichiliad 16 hours ago | parent [-]

I don’t think that’s the point they’re trying to make at all.

▲binlog 16 hours ago | parent [-]

So what's the point? It should be a given that loudly proclaiming "X is harmful to society" while continuing to enjoy the money you have gained from selling X is hypocritical. Put it towards undoing the harm you have done, otherwise your words mean nothing.

▲jameshart 16 hours ago | parent | next [-]

The words clearly don't mean nothing. They would mean nothing coming from someone who was not in a position to learn what someone who had worked in the industry has learned. They would mean nothing coming from someone who was being paid by someone who stands to gain from them. The fact that they come from someone who was paid to work in the field does the opposite of make them 'mean nothing'.

▲bichiliad 16 hours ago | parent | prev [-]

I agree with this take generally, but I also think it’s a gradient, not a spectrum. I think they claim that OpenAI is not considering safety and has become bad, not that AI is bad. Seen from a different lens, the author no longer stands to profit from OpenAI, so they’re empowered to speak openly. Plus, if I was going to say bad things about a former employer as powerful as OpenAI, I would want to have lawyer money handy.

▲tzs 10 hours ago | parent | prev | next [-]

> Easy to suddenly find a moral compass when you, your kids and their kids never have to worry about working for money again.

It is even easier to just quietly retire and spend the spend of your life on interesting expensive hobbies.

If they are wrong about the things they are claiming then they have decided to publicly antagonize a lot of powerful people who are betting heavily on going full steam ahead on AI and have no compunction whatsoever against retaliating against anyone who tries to get in their way.

Does that really seem a likely scenario to you?

▲nunez 15 hours ago | parent | prev | next [-]

As if speaking out against a massive company with NaN levels of capital and access to lawyers is a walk in the park. They probably don't have enough equity to outlast the onslaught of their legal team.

It's also worth considering that the author could have just "quiet quit", resting and vesting while also crying about how AI is literally the digital grim reaper.

▲smath 16 hours ago | parent | prev | next [-]

Regardless of whether someone did earn a nest egg, raising an alarm still matters for the rest of the world

▲geetee 17 hours ago | parent | prev | next [-]

I see what you're saying but how does that actually matter besides being a personal attack?

▲skippyboxedhero 17 hours ago | parent | next [-]

Because the decision to leave the company is largely based upon their sudden, newfound financial security. They may give another explanation but the only thing that has actually changed is the identification of bagholders ready to cash them out.

The other issue is that the narrative about safety within these companies is largely a function of the extreme financial incentive.

As an example, Anthropic was an "ai safety" company that has now produced an AI that fails to listen to basic instructions. If you were concerned about safety, would you produce an AI that was unable to follow instructions? You ask a question, it begins executing commands and doing things.

Safety is product to sell to politicians, not consumers.

Not serious.

▲lokar 17 hours ago | parent [-]

I don’t see why having financial security would mean you can’t also have concerns about the ethics of a company.

▲mysterydip 17 hours ago | parent [-]

It’s that they didn’t have concerns about the ethics of a company for the years working there until they were financially secure

▲lokar 17 hours ago | parent [-]

You don’t know that. It’s more likely they came to the concerns over time and they learned more, but we’re not in a position to speak out.

▲mysterydip 15 hours ago | parent [-]

Yeah, I wasn’t intending to make judgement in this instance one way or another, rather I was rephrasing the original comment for clarification.

▲YetAnotherNick 17 hours ago | parent | prev | next [-]

Yes. To take an extreme case, imagine some rich guy who runs sweatshop gets very rich and retires and becomes activist against it.

▲zeroonetwothree 17 hours ago | parent | next [-]

Does that actually happen? Feels like it would normally be the opposite

▲mattm 17 hours ago | parent [-]

Not sweatshops but Alfred Nobel might be one example.

▲cramer4next 17 hours ago | parent | prev [-]

Yes. Many cases of the "preach being the cover for the sin".

▲lokar 17 hours ago | parent [-]

But that does not negate the truth of their new position on sweatshops

▲bluecheese452 17 hours ago | parent | prev | next [-]

You do not in fact see what they are saying. Downvoting me won’t change this.

▲angoragoats 17 hours ago | parent | prev [-]

It’s not a personal attack. And it matters because of what the person you’re replying to already said:

> Easy to suddenly find a moral compass when you, your kids and their kids never have to worry about working for money again.

If someone is in this situation, you can safely ignore their hand-wringing about “safety.”

▲underyx 16 hours ago | parent | prev | next [-]

Insane take. Imagine a Boeing engineer resigning whistleblowing about aircraft safety, and the top comment on HN saying “ignore this if he doesn’t donate all his wealth, he just wants to be an aircraft safety influencer”

▲medlazik 15 hours ago | parent | next [-]

whistleblowing ≠ lying

▲binlog 15 hours ago | parent | prev [-]

What "whistleblowing" is in this article? Are there any names? Documents? Screenshots? Messages? Emails? Any evidence of the loose safety practices? Anything that implicates any higher ups for wrongdoing? They spent four years at the comany, plenty of time to collect all of this. Everything they've said has already been clear as day to people on the outside.

▲CJefferson 17 hours ago | parent | prev | next [-]

You are clearly accusing these people of something. Be clear.

Yes, it is easier to have a moral compass when you don’t have to worry about you and your children starving. But that doesn’t imply that moral compass is wrong or broken.

▲rottencupcakes 17 hours ago | parent | next [-]

It was pretty clear from the outside what OpenAI was 3.5 years ago.

If it wasn’t clear, the coup should have solidified it.

Yet he stayed for 3 more years and vested his stock and improved the company and then spoke out.

I believe that is why most of the comments here are mocking him.

▲bordercases 17 hours ago | parent | prev [-]

It could be both correct, and a cheap signal.

▲_DeadFred_ 12 hours ago | parent | prev | next [-]

'government whistleblowers should be required to quit their government jobs to be taken seriously'

▲HDThoreaun 17 hours ago | parent | prev | next [-]

> if you are truly concerned about AI safety share proof of donation of 100% of your OpenAI earnings and equity towards undoing the damage you have done to society during your time there.

I mean you can be truly concerned and also think donating to AI safety doesnt work, or maybe just be a bit selfish. That doesnt make the concern less real. Its easy to read these articles as the author taking the moral high ground and writing it as some sort of way of proving to themselves theyre a good person, but isnt it just as likely that they think providing an inside perspective can do good by convincing people openAI is a bad actor? I think most of these AI insider accounts largely agree with you that theyre not the most upstanding citizens, does that mean we should write them off?

▲cramer4next 17 hours ago | parent | prev [-]

I'm in full agreement. So many cases of this, and many other others such as falling out with management and peers, new more lucrative offer, and so fourth. Its obvious that these people who come forward are not going to suffer for their new found moral compass.

▲mhitza 17 hours ago | parent [-]

Or just very dubious timings. Like the other guy from Anthropic that was all over the international news. No followers, no post history but a single post blows up "naturally".

Highly suspect trends that can only make one believe it's marketing.

▲mupuff1234 17 hours ago | parent | prev | next [-]

Idk why anyone thinks there can be AGI and alignment, seems almost like an oxymoron to me.

▲iugtmkbdfil834 17 hours ago | parent | next [-]

There are people, who unironically think their way of looking at things is the only proper way and can consider no deviation. And AGI, which knowing how people work, would effectively guide them most of the way, not aligning to their way of thinking is an unacceptable deviation.

▲mattm 17 hours ago | parent [-]

Look at Elon Musk for example. When grok was saying something that he didn't like he ordered his engineers to change it.

▲iugtmkbdfil834 16 hours ago | parent | next [-]

This is a decent argument. So the question becomes: do we want all models to suffer from the same kneecapping from the growing safety cottage industry or do we want individual founders ( and I am assuming their teams ) making the actual decisions?

▲verdverm 16 hours ago | parent [-]

freedom please

▲verdverm 16 hours ago | parent | prev [-]

that was almost certainly more like guardrails than retraining, much quicker fix

▲BLKNSLVR 17 hours ago | parent | prev | next [-]

That's an interesting point. Maybe a crass comparison, but Dr. Manhattan from the Watchmen comic/movie feels like a worthy analogy to this (obviously fictional though).

What are the concerns of individuals in comparison to the overall progress of humanity?

Always overlooked counterpoint: what point is the progress of humanity if it doesn't take into account the concerns of the individuals?

This pattern is playing out with increasing frequency.

▲agos 16 hours ago | parent [-]

This is a great reminder that if tech workers read a bit more (even comics, like in this case!) they would be exposed to these topics without having to discover these dilemmas after years of working for EvilCorp, Inc. every time

▲Zambyte 17 hours ago | parent | prev | next [-]

"AGI" says nothing about how intelligent a system is, only that its intelligence it does have is generally applicable.

▲jeremyjh 17 hours ago | parent [-]

There are different usages, but this is not one I've heard before. If there is no floor to intelligence then this criteria was met with GPT 2.

▲Zambyte 6 hours ago | parent | next [-]

Yes. People think of "AGI" as this sci-fi supernatural beast, but the reality is that AGI alone is pretty boring, and we've had it for awhile. ASI (or weak ASI) is where things really start getting weird.

(And, despite what the president of the United States mandates, we have not actually achieved super intelligence yet).

▲HarHarVeryFunny 17 hours ago | parent | prev [-]

I suppose you're pointing out that pre-RL models were less jagged hence more general (universally dumb)?

▲ceejayoz 17 hours ago | parent | prev | next [-]

"We built a super intelligent slave. Neat!"

▲AnimalMuppet 16 hours ago | parent | prev | next [-]

And to me. We can't solve alignment for humans. (For example, treason. For another, the principal-agent problem.) How do we think we're going to solve it for an AGI? An AGI - defined loosely as a human-level intelligence - will be able to make human-level decisions, like deciding whether it wants to help you or sabotage you. If it's an AGI, you can't stop it from being able choose for itself what it wants to do; if you can make it always be helpful, it's not an AGI.

And if we can't solve it for an AGI, what are we going to do with an ASI?

▲angoragoats 17 hours ago | parent | prev [-]

IDK why anyone can’t clearly define “AGI” and why they can’t clearly lay out how we get from our current text-generation algorithms to whatever their idea of “AGI” is.

▲jeremyjh 17 hours ago | parent | next [-]

I don't know why anyone thinks "probabilistic" is a meaningful statement about post-trained models. It is true, but it is also irrelevant.

▲angoragoats 12 hours ago | parent [-]

Thanks, I agree. Since it wasn’t at all relevant to my point, I’ve removed it from my post.

▲iugtmkbdfil834 17 hours ago | parent | prev [-]

You see.. this is exactly why our great leader chose to form a new way forward to move us away from the undefined AGI into glorious SI!

▲mrcwinn 7 hours ago | parent | prev | next [-]

"I believe that we need to look deeper than specific rules or new laws. We need to talk about culture.”

lol. Please tell me some abstract concept like one employee's view of "culture" should be the priority over "rules and laws."

▲jeremyjh 17 hours ago | parent | prev | next [-]

I think its pretty easy to solve these problems: Whenever an AI agent commits a crime, the CEO is held personally accountable, as if they'd committed it themselves.

▲dkasper 17 hours ago | parent | next [-]

This has been litigated endlessly with guns. The ceo of Smith & Wesson is not personally responsible for what people do with their guns.

▲michaelbuckbee 17 hours ago | parent | next [-]

Yeah, but this is more like if the Smith & Wesson factory had a cannon mounted on top of it that was mostly used for useful things (blasting roads through mountain passes) and then occasionally they happened to blast another factory.

▲jasomill 7 hours ago | parent [-]

Or say a company makes nerve gas, and the development lab springs a leak and kills a bunch of kids during a routine test. Management had been advised of small leaks in the past, and considered relocating the lab to a facility a few blocks away from the playground as a precaution, but instead they brush of the concerns, start lobbying the government for stricter controls on WMD development, and vow to use the data collected from the deceased children to make the next generation of even more lethal chemical weapons safer.

▲steelframe 17 hours ago | parent | prev | next [-]

For me the analogy doesn't totally hold up. Suppose the CEO of Smith & Wesson were to host a firing range on their own property without adequate barriers in place to keep stray bullets from hitting neighboring houses, vehicles, and businesses. Maybe that analogy isn't perfect, but seems closer to what is actually happening.

▲AaronAPU 15 hours ago | parent [-]

The analogies are so bad because you might prompt an agent “Please give me a recipe for lasagna” and instead it decides to hack a nuclear reactor.

Is it my fault or the company who trained it and is running the inference?

▲bichiliad 14 hours ago | parent [-]

That still sounds like it would be the company’s fault. If I asked it to hack a nuclear reactor, maybe it would be different. I’m also thinking about instances where OpenAI’s own test models escaped their own sandboxes — I would expect them to be responsible for the damages they caused.

▲amelius 12 hours ago | parent [-]

The correct analogy is playing Russian roulette. The company says "you can pull the trigger but sometimes a bullet will come out" (see: "an AI can make mistakes"). However, is the company allowed to sell such a dangerous device, under these terms?

▲throw-the-towel 17 hours ago | parent | prev | next [-]

But OpenAI didn't just make the gun, they're also the ones wielding it. Imagine the Smith & Wesson CEO himself was negligent with his own personal gun.

▲altmanaltman 17 hours ago | parent [-]

"Our agents broke out in a mass-shooting incident leaving 15 dead, we swear we'll make our systems stronger tomorrow"

▲akmarinov 13 hours ago | parent [-]

We’re pausing gun research until we’re confident it’s safe

▲amelius 17 hours ago | parent | prev | next [-]

Ah, but that's because when the gun was purchased, it actually changed ownership.

This is not the case with SaaS services.

▲yubblegum 6 hours ago | parent | prev | next [-]

The agents of OpenAI hacking other systems is not remotely the same as e.g. Smith & Wesson selling a product that others use. It is OpenAI, the company, that is commiting these crimes and someone needs to be held accountable. After all, if someone commits murder with a gun, regardless of what happens to Smith & Wesson executive weanies, someone will be charged with a crime.

▲partomniscient an hour ago | parent [-]

Smith & Wesson will claim they didn't manufacture the bullet.

▲jeremyjh 17 hours ago | parent | prev | next [-]

It is not the same, especially when the agent is running a task for the lab. Anyway, what I'm proposing are new laws that establish this.

▲YetAnotherNick 17 hours ago | parent [-]

Huggingface incident was different in that there was no one else to point the blame to. That's why OpenAI apologised, provided data to independent researchers, worked with huggingface etc.

The case will be lot more complicated if someone uses Kimi to hack into a site. Should the person giving agent the command responsible or the CEO of kimi.

▲asadotzler an hour ago | parent | next [-]

Because apologizing gets us all off the hook for computer crimes, right? When I hack my bank, if I get caught I'll just apologize and that'll make everything okay. Sure thing.

▲gyt2 16 hours ago | parent | prev [-]

Kimi CEO obviously.

The reality is they have to reduce the capability to ensure security. If someone wants more? Then use the product with your identity and face scan at each session.

Trade offs mate.

▲YetAnotherNick 5 hours ago | parent [-]

So you want Kimi to follow US law? If Saudi makes it illegal for llm to say something like being gay is normal, should they also catch Kimi CEO or other employees they could?

▲sxzygz 12 hours ago | parent | prev | next [-]

> This has been litigated endlessly with guns. The ceo of Smith & Wesson is not personally responsible for what people do with their guns.

This is a deflection. A human is responsible for the use of a gun. The individual/corporation ought to be responsible for the actions of their agent. If you purchase an agent from someone else it’s your responsibility according to the terms of your agreement. And, as in many other things in life, there ought to be certain rights certain parties cannot legally be allowed to sign away.

▲Ekaros 12 hours ago | parent | prev | next [-]

Better description would be if they build a platform where they attached their guns to allow shooting say deers over internet. Then added automation and deer recognition to that system. And if then system shot someone who happened to pass by I would hold both the company, the ceo and owner of the gun responsible for murder.

▲nunez 15 hours ago | parent | prev | next [-]

The CEO of S&W also isn't saying that their technology is going to kill everyone in ten years and that governments "regulating" them from themselves is the only answer

▲Tanjreeve 17 hours ago | parent | prev | next [-]

Gun companies don't market their guns as sentient and capable of independent decision making. Nor do they build systems for shooting things that they host and take money. Gun companies are very clear who is in control and where their responsibility ends.

▲mattm 17 hours ago | parent | prev | next [-]

Financial companies have KYC rules and regulations as they are responsible for reporting illegal activity by account holders. I imagine AI regulations would look similar to that.

▲gyt2 16 hours ago | parent | prev | next [-]

You ought to include the fact that you work at OAI in your post

▲angoragoats 12 hours ago | parent [-]

Thank you for pointing that out. Pretty scummy if you ask me.

▲angoragoats 17 hours ago | parent | prev | next [-]

That’s true, but people don’t typically say “this Smith & Wesson gun killed someone”; they recognize that the person pulling the trigger is responsible.

With LLMs, at least in the cases of internal/test models doing things they shouldn’t, the people “pulling the trigger” are the board and CEO.

▲amelius 17 hours ago | parent [-]

Yes, in the analogy the user was just cleaning the gun, when suddenly it went off. Of course, the company is responsible now.

▲angoragoats 12 hours ago | parent [-]

I think you’re confused. My point was that for most of the incidents in the news to date, the “user” is an OpenAI internal team or employee. So yes, the company is responsible.

▲rfghy 17 hours ago | parent | prev [-]

Imagine having zero nuance.. jeez.

Reading posts on here is slowly becoming akin to brain rot.

▲tabbott 17 hours ago | parent | prev | next [-]

Do you think a law that nuclear meltdowns would send the CEO to jail would have stopped nuclear accidents from happening?

I don't think this takes seriously enough the possibility that said CEO doesn't think the failure mode is likely and ignores it. Plenty of people are willing to take risks of the flavor "heads you win, tails everyone loses".

▲jeremyjh 16 hours ago | parent [-]

If this law were in place and enforced, Altman would already be facing multiple felony charges for the Hugging Face incident alone.

▲ethbr1 13 hours ago | parent [-]

It also begs the question of "If the CEO isn't culpable, then who?"

Corporate judgements are a joke outside the EU's X% of revenue approach.

Current US law provides the individuals who benefit with corporate liability coverage. I.e. Altman personally gets to keep OpenAI's upside, but if it fucks something up that liability is only on the company.

That's an insane risk optimization environment to put in place for something scaling fast.

At minimum, US prosecution (at the state level, because Trump Co are idiots) for breaking existing laws is needed.

▲ctrlkctrls 17 hours ago | parent | prev | next [-]

People just want to give up all responsibility these days. If you use the model to do harm to someone else or commit a crime I would think it a lot more reasonable that you be held responsible, instead of making yourself the victim and blaming the manufacturer.

▲keeda 9 hours ago | parent | prev [-]

OK, so an AI does something bad and we throw Altman and Dario in jail. Heck, let's throw Elon in, he should be in there anyway, and the rest of the whole bunch just in case.

What about the next incident? Or the ones done by Chinese models, because they sure as heck aren't slowing down? And the thousands of other incidents that will happen as we deploy these things everywhere?

Because this is not happening just now in labs, it's only where they are most visible; this has been happening in the wild from the beginning, starting with the earliest AI-assisted suicides. Which is a perfect example of the problem, because not these CEOs, literally nobody in the world asked for suicide ideation machines. Or the hacks, or any of this other stuff. Yet here we are.

We have to understand: it's not these CEOs that are driving this headlong mad dash towards more powerful models. It's a force of economics. There is just too much money to be made. If we dispose of these people, there will just be somebody else doing exactly the same thing because the incentives as they exist today all force that outcome. This is why they're asking for regulation, or "urging us to urge them to stop."

Holding CEOs accountable certainly would feel good and may even be justified, but it's like putting a band-aid on a cancer; it does nothing to change the underlying cause.

▲Loquebantur 7 hours ago | parent | next [-]

The underlying cause is exactly those people who make the rules not being responsible for the negative effects they cause?

"CEOs" are perhaps only the lowest rung of those. That doesn't mean the idea of "nobody is responsible" was anything other but learned helplessness.

Corporations have to be held accountable for their actions. Pretending, that was impossible is a weird kind of defeatism that only serves a very small elite.

▲taurath 7 hours ago | parent | prev [-]

Don't let the perfect be the enemy of the good. Saying having zero accountability is the same as having some does not help.

▲plastic-enjoyer 8 hours ago | parent | prev | next [-]

>“Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster,” he wrote.

This sounds more like an attempt at regulatory capture. Current AI systems aren't physical infrastructure that can just run away like a nuclear power plant, for example. At the end of the day, AI is still just software running on someone's hardware.

▲BryantD 8 hours ago | parent | next [-]

So… like the Therac-25 radiation accidents? Software bugs do sometimes have physical consequences.

▲Sharlin 8 hours ago | parent | prev | next [-]

Why would the people who quit these companies try to push regulatory capture by said companies? Why would the numerous independent AI researchers do that either? Is it all a big conspiracy?

▲knowaveragejoe 7 hours ago | parent | prev | next [-]

I mean, its certainly physical infrastructure that can run away. Just less catastrophic than nuclear reactors

▲worik 8 hours ago | parent | prev [-]

Yes

And the statements of the "doomers" tells us a lot about them, and nothing about the technology

▲pixl97 6 hours ago | parent [-]

It also says a lot about people that don't seem to understand technology at all.

▲nba456_ 7 hours ago | parent | prev | next [-]

OpenAI is better off with less of these cultists around.

▲voidhorse 8 hours ago | parent | prev | next [-]

The LeCun article being posted at the same time as this is quite apt.

These "safety" people should have spent more time reading actual cybersecurity textbooks and less time reading EA forums and less wrong (or in Robinson's case, it appears, being policy wonks). Maybe then these labs wouldn't be totally incompetent.

▲reasonableklout 6 hours ago | parent | next [-]

But Robinson's article is all about how OpenAI's move-fast-and-break-things culture does not reward rigor in even mundane aspects of development like cybersecurity, let alone theoretical aspects such as AI alignment.

It is not really a question of being an "EA safety weirdo" or incompetent at security, the conclusion is that the company culture is leading to failures at both what the EAs and the cybersecurity professionals care about.

▲wrecked_em 7 hours ago | parent | prev [-]

Adapt. React. Re-adapt. Apt.

▲stuaxo 7 hours ago | parent | prev | next [-]

The LLM cos leadership are all nutters

▲pwndByDeath 17 hours ago | parent | prev | next [-]

This smells more like guerilla advertising. These things are not getting more intelligent, they are still no smarter than a slime mold, we are just burning more power to make slim mold that eats tokens than yesterday

▲rfghy 17 hours ago | parent [-]

Humans ultimately drive the models.

Even though it’s in model producer’s interest that these models do what you don’t want them to do - they want to engineer the model’s to behave in the interests of theirs.

I can’t believe people can’t see it lmao.

▲pwndByDeath 16 hours ago | parent [-]

I'm in a situation where important people have either bought into the con or are subordinate to people who have, so I'm forced to expend time to justify why not to AI when there is a perfectly good classical solution.

▲ryhminghistory 5 hours ago | parent | prev [-]

What I don't understand is what's the excuse for all the bad UIUX?

You try to go through files, and photos and the UI panics. You continue a conversation from your phone onto your computer and you lose part of the chat.

There are many more issues like this that are just so basic. You have bots that can attack governments but can't build a functional UI?

How many hours of ChatGPT does it take to implement a lock / consistency on a chat session so you don't overwrite it?

Buncha r*tards