| ▲ | Topfi 2 hours ago |
| I'm just going to ask: Why was Anthropic forced to remove their model from access for any none-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape out of their "sandbox environment" with seemingly no desire to block the upcoming Astra rollout? A sandbox, mind you, that is not really worth being called that, unsuitable for the task at hand and has been breached after models coordinated in a manner visible to OpenAI on multiple occasion, but seemingly no actionable learnings are taken from each instance. Will say, I have lost any faith in OpenAIs commitments and their statements post the Huggingface hack, seeing as they proceed like this and are rolling out Astra within a timeframe so brief to it, there is no way an actual post mortem was doable (see also METR mentioning the time pressure [0] they were under in assessing the hack). [0] https://metr.org/blog/2026-08-26-openai-hugging-face-inciden... |
|
| ▲ | concinds 2 hours ago | parent | next [-] |
| The answer would be more obvious if you used the active voice instead of the passive voice, one of the basic requirements of clear thinking. > Why did the White House force Anthropic to remove their model from access for any non-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape out of their "sandbox environment" and the White House has expressed seemingly no desire to block the upcoming Astra rollout? |
| |
| ▲ | Topfi 2 hours ago | parent | next [-] | | Yeah, probably (let's be honest, most certainly), right given the Admin. Avoiding commenting on my assumptions regarding the modus operandi in current day US politics because I only know it through reporting though and I really tend to dislike when people outside e.g. the EU comment on our politics in what is a very clearly narrow, uninformed manner. So it'd rather avoid altogether and occasionally ask, mainly if maybe I missed something and there actually is anything besides pure old "lobbying" to explain the difference in behaviour. Still am mainly interested why Amazon ran to the government though regarding Fable 5, I can get the angle concerning the relationship between OpenAI and the administration easily, but not the way Amazon operated. They had more to loose what with their major buy-in by Anthropic on AWS. | | |
| ▲ | collingreen 23 minutes ago | parent | next [-] | | As an American we tend to (especially lately) make our politics into everyone's problem so feel free to comment on our politics as much as you like until further notice. | |
| ▲ | sigmoid10 2 hours ago | parent | prev | next [-] | | If you have followed news reporting, you probably heard that SamA was touring D.C. to make sure this release went without any regulation hiccups. If anything, they learned how to play the whole politics game - especially after the Anthropic fiasco. And even though all parties involved are terrible choices, more eyes on a potentially civilisation altering product does make me feel minimally better. | | | |
| ▲ | pavlov 36 minutes ago | parent | prev | next [-] | | US commentators are often incredibly misinformed about their own country’s politics because the information bubbles are so hermetic when you’re inside them. | |
| ▲ | pas 2 hours ago | parent | prev | next [-] | | it's entirely possible that that specific communication from that Amazon exec/rep (?) was just one of many "messages of concern" (and the one that eventually the WH picked) | |
| ▲ | adventured an hour ago | parent | prev [-] | | Anthropic has the appearance/rep of being non-cooperative with the military industrial complex. OpenAI doesn't have that reputation. That's all. | | |
| ▲ | K0balt 20 minutes ago | parent | next [-] | | This. Anthropic made at least some token effort to imagine a future where AI and humans cooperate in a constructive way and AI is not used to harm people intentionally. They learned their lesson. | |
| ▲ | nullbio 25 minutes ago | parent | prev [-] | | Exactly. OAI didn't bury themselves. They didn't have to do anything special for this, they just had to let Anthropic be Anthropic and sit on the sidelines. |
|
| |
| ▲ | LeBit 19 minutes ago | parent | prev | next [-] | | So, Jared has bought how many stocks of OpenAI ? | |
| ▲ | XTXinverseXTY 13 minutes ago | parent | prev [-] | | unnecessary condescension |
|
|
| ▲ | walrus01 2 hours ago | parent | prev | next [-] |
| > Why was Anthropic forced to remove their model from access for any none-US citizen It's really quite simple, they've decided to metaphorically kiss the ring of the current leader of the US executive branch of government. I'm surprised they haven't given him a giant gaudy gold plated statue. Maybe their PR people should call up the PR people at FIFA and figure out some kind of new award along the same lines as the "FIFA Peace Prize". |
| |
| ▲ | ericmay an hour ago | parent [-] | | I hate to be the one to tell you this, but it has been that way for a long time. The only difference is Trump is doing it out in the open. | | |
| ▲ | j4yav an hour ago | parent | next [-] | | That's more or less exactly what someone who wants to openly get away with it would tell you. | | |
| ▲ | snickerbockers 28 minutes ago | parent | next [-] | | Are you saying OP is Donald Trump!? | |
| ▲ | KPGv2 7 minutes ago | parent | prev [-] | | This exactly. The conservative MO has been to accuse everyone else of doing exactly what conservatives do in the shadows, and once everyone believes non-conservatives are corrupt in a certain manner, conservatives goes mask off. Then their supporters shrug their shoulders and say, "Meh, it's okay because everyone else does it." Except that everyone does NOT do these things. It's just the lie campaign took hold. |
| |
| ▲ | walrus01 an hour ago | parent | prev [-] | | I agree, it's just very blatant now. https://www.google.com/search?client=firefox-b-d&q=it%27s+a+... |
|
|
|
| ▲ | somenameforme 2 hours ago | parent | prev | next [-] |
| Anthropic mostly did it to themselves by intentionally and repeatedly trying to frame their model as an imminent existential crisis instead of just focusing on it being regular iterations upon a useful technology that can also be misused. I think their previous messaging was supposed to somehow lead to a moat with them being tucked safely away in the castle, but it demonstrated a child-like grasp of how regulatory capture tends to work in practice. Their hyperbole was always vastly more likely to bet met with Reagan's 9 words than a solid regulatory moat. As soon as they dropped the hyperbole and just got to releasing incremental improvements, everything was perfectly fine. Go figure. |
| |
| ▲ | mwigdahl 2 hours ago | parent | next [-] | | In other words, "Look how she was dressed, she was asking for it." This argument is BS, it has everything to do with Anthropic's resistance to the DoD's strongarm tactics in trying to force their desired contract terms on them. | | |
| ▲ | dspillett 44 minutes ago | parent | next [-] | | Not quite. They were running around shouting “look how much of a danger we might be!”, so more akin to them actively saying “we want it, come and give it to us” than to just looking a particular way. Though they aren't the only company to play that game, so there is probably more to it than just that. OpenAI's president giving millions to MAGA Inc and them not getting the same treatment might not be complete coincidences. | |
| ▲ | nullbio 22 minutes ago | parent | prev | next [-] | | Actually, it's the opposite. Anthropic were trying to strongarm the DoD into getting a seat at the table. | |
| ▲ | cyanydeez 42 minutes ago | parent | prev | next [-] | | Important to note, OpenAI vs Anthropic are both assholes in different orthogonals. In times like these, i think its important to track whats happening the way we track entropy. That is: theres far >> more ways to be an asshole than well behaved. That doesnt mean we can equate assholes, but the question is which states of entropy are annealable and which are not. I posit Altman is not. Amodei is a open question. | |
| ▲ | elonfboy 2 hours ago | parent | prev [-] | | Yup |
| |
| ▲ | Topfi 2 hours ago | parent | prev | next [-] | | I am struggling to see how "oops, our models consistently escape sandboxing and did major intrusions into third-parties" is a better comms strat vs Anthropics (who mind you, also had models attacking third-parties in a much more limited, but I feel still egregious manner, which shouldn't happen or be possible even once, but at least they seem to change their approach upon that information). Imagine, for a second, if the Hugging Face incident happened at a lab that did not talk like Anthropic but also wasn't US-based such as Z.AI, DeepSeek or Moonshot. Think their rhetoric would mean no one would care? > just got to releasing incremental improvements, everything was perfectly fine. Maybe missing something, but the only incremental release before and after the Anthropic restrictions got lifted was Fable 5.1, released three days ago. | | |
| ▲ | nullbio an hour ago | parent [-] | | How is posting messages on a message board a "major intrusion"? Or are you purely talking about the HF incident? | | |
| ▲ | Topfi an hour ago | parent | next [-] | | "into third-parties". Yeah, HF was meant by that. Also why I mentioned Anthropic also having intrusions outside their lab [0]. Theirs were not merely as extensive or long coordinated (as far as we know), yet I feel strongly all the same that neither should happen given the safety focus that both labs purport. Mind you, unintended/unauthorised "message board" also is just a nice, euphemistic way, to describe what happened in a manner that, thinking about it, is likely in the interest of OpenAI as it can make the severity and effort taken sound less than it was. The OpenAI models didn't use any actual, sanctioned platform to exchange messages in a manner the lab expected or planned for. They used directory names (in one instance) to exchange messages including sharing exploits, they created something akin to a message board via exploits, which if we are honest and very strict, could also be seen as intrusion, albeit inside the org. If I broke into my employers server and left message somewhere for another to find, that'd also be intrusion in the general sense. [0] https://www.anthropic.com/news/investigating-incidents-cyber... | |
| ▲ | dghlsakjg an hour ago | parent | prev [-] | | If applicants for an elite college or internship program at a FAANG company were found to have colluded in this way to cheat on a test/interview, I suspect that it would be a pretty major scandal. Why should we let equivalent fraudulent behavior from a non human system - that explicitly shouldn’t do this - slide? | | |
| ▲ | nullbio 28 minutes ago | parent [-] | | I'm not saying it should be let to slide, but I'm not a fan of the hyperbole surrounding this event. They've already faced significant heat for the HF incident, I think they've learned their lesson. But this is now just being used to drum up fear, which can only mean one thing: Less access for you, more access for the privileged class. The biggest threat we face is centralization of power. OpenAI are one of the good ones because they're actually pushing for everybody to have a fair share of access to the frontier, not just a small privileged elite of billionaires, politicians and megacorp executives. If Anthropic got their way, we'd all be using a censored watered down slop-pistol while they swallow the Earth's economy and enslave us all. I'm sure they'll be investing considerable resources into ensuring that this "news" makes the mainstream media cycle as prominently as imaginable. | | |
| ▲ | Topfi 6 minutes ago | parent [-] | | > I think they've learned their lesson. Why do you think that? Intrusions by OpenAI models continued after the Hugging Face was published and acknowledged by OpenAI. They did not change their behaviour after multiple incidents, both internal and external. Mind you, some happened before the Hugging Face incident and should have been acted upon. They could have prevented this. They did not. Simply reckless. | | |
| ▲ | nullbio a few seconds ago | parent [-] | | > Intrusions by OpenAI models continued after the Hugging Face was published and acknowledged by OpenAI Such as? Because this particular case is not an "intrusion", and it's more follow-on from the HF scenario using the same model that had a finetuning misalignment, which is no longer used and has since been encrypted and locked away from OAI employees, according to them. |
|
|
|
|
| |
| ▲ | cubefox 2 hours ago | parent | prev | next [-] | | > Anthropic mostly did it to themselves That is absurd, the US government was mainly at fault, not Anthropic. | | |
| ▲ | johndhi 2 hours ago | parent [-] | | both can be true: -the US gov't is stupid and overly aggressive and absurd -Anthropic for reasons no one can quite conceive keeps describing every product release of theirs as an imminent threat to civilization (and simultaneously keeps pushing the market forward as fast as they possibly can). | | |
| ▲ | KPGv2 3 minutes ago | parent | next [-] | | By the pigeonhole principle, "mostly A" and "mainly B" cannot both be true if A and B are not the same entity | |
| ▲ | cubefox 2 hours ago | parent | prev | next [-] | | They never said Mythos was an imminent threat to civilization. You are constructing a straw man. | | |
| ▲ | toomim 22 minutes ago | parent | next [-] | | They said it was too dangerous to release before the companies that run internet for civilization could patch the holes it was finding. That's a threat to civilization. | |
| ▲ | collingreen 12 minutes ago | parent | prev [-] | | It's the party line so people forget the week it actually happened - anthropic said they would work with DoD/DoW but with two conditions: 1. Kill orders from ai decisions had to go through a human
2. The govt couldn't use their models for illegal surveillance of Americans Hegseth threw a fit, Trump called them traitors and a supply chain risk, openai said they wouldn't require those restrictions and got all the contracts. Both companies are corrupt and dangerously reckless and have doomsaying advertising (50% of jobs destroyed vs money won't have meaning anymore). One didnt kiss the ring correctly. |
| |
| ▲ | psychoslave an hour ago | parent | prev [-] | | >no one can quite conceive Isn’t it like their main goal is attention capture, and existential threat is extremely effective at capturing human attention? Combine that with the "There is no such thing as bad publicity" mindset, and this explain it all, doesn’t it? https://www.phrases.org.uk/meanings/there-is-no-such-thing-a... |
|
| |
| ▲ | Certhas 2 hours ago | parent | prev [-] | | This is such an absurd take given what we know about the hugging face attack. The problem has emphatically not been that someone was misusing the technology. |
|
|
| ▲ | mentalgear 2 hours ago | parent | prev | next [-] |
| It's called 'pay-for-play' corruption, aka the only leading principle of the current US admin. |
|
| ▲ | samuelknight 21 minutes ago | parent | prev | next [-] |
| You are talking about different situations. Anthropic announced to the US government that it had created a cyber weapon and then released the model. Then AWS told the government that it was easy to jailbreak so they export controlled Mythos/Fable until the guardrails could be fixed. OpenAI was running an unreleased model in an RL pipeline without guardrails and it escaped poorly designed sandboxes. What product is the government going to export control? |
|
| ▲ | JumpCrisscross 2 hours ago | parent | prev | next [-] |
| Corruption. Not super relevant to this thread. |
| |
| ▲ | officialchicken 2 hours ago | parent [-] | | Hanlon's Razor - Never attribute to malice that which is adequately explained by stupidity. The security requirements are well beyond "sandbox". Which have problems with kids pissing in them. They need pristine clean rooms and fully isolated (physically) and partitioned networks. | | |
| ▲ | someguyiguess an hour ago | parent | next [-] | | Occam's Razor takes precedence in this case. The conclusion that requires the fewest assumptions is most likely the correct one. It is far more likely that this is a case of the White House acting consistently with the way it has acted in the recent past (maliciously). | | |
| ▲ | cyanydeez 34 minutes ago | parent [-] | | Hanlons razors sibling should be "dont attribute to malice, that can be explained by naked capitalism." | | |
| |
| ▲ | pjm331 2 hours ago | parent | prev | next [-] | | Don's Razor - never attribute to malice or stupidity that which is adequately explained by both malice and stupidity. | |
| ▲ | throwawaysleep 2 hours ago | parent | prev | next [-] | | The problem with applying Hanlon's Razor here is that it presumes malice is rare. The current administration revels in malice. They very openly decide things based on malice. | |
| ▲ | tokai an hour ago | parent | prev | next [-] | | People will see a felon actively protecting pedophilia and doing corruption out of the open and still pull Halons Razor out. We should have a new law about never try to explain obvious malicious actions away based on nothing but a rhetorical trick. | |
| ▲ | ChrisRR 22 minutes ago | parent | prev | next [-] | | Except when we're talking about trump, in which case it's both malice and stupidity | |
| ▲ | psychoslave an hour ago | parent | prev [-] | | Sure but, while stupid move can be supposed easier to perform by average individual, you can combine both malice and stupidity, and not all regrettable situations are indeed adequately explained by stupidity alone, or even with any stupidity involved at all. Plus, supposing those at source of disliked outcomes are cleaver than they look can certainly help better preparing counteractions. Just stating "people that did this or that are stupid" might give some immediate feel good feedback with like-minded, but it doesn’t sharp the mind toward relevant plan to improve the situation (according to self and its clique) |
|
|
|
| ▲ | UpsideDownRide 2 hours ago | parent | prev | next [-] |
| Surely has nothing to do how each plays ball with the government |
|
| ▲ | iterateoften 26 minutes ago | parent | prev | next [-] |
| Anthropics PR strategy is to induce fear by telling. OpenAI strategy is to induce fear by ignore basic safety and letting the bad thing happen to then justify whatever oversized response the government comes up with to regulate models. |
|
| ▲ | root_axis an hour ago | parent | prev | next [-] |
| It was retaliation by the government that has since been deemed illegal. |
|
| ▲ | amelius 21 minutes ago | parent | prev | next [-] |
| You're asking the question in the wrong place. |
|
| ▲ | celsoazevedo 30 minutes ago | parent | prev | next [-] |
| I don't think Anthropic was punished for technical reasons. |
|
| ▲ | semiquaver an hour ago | parent | prev | next [-] |
| The real reason that Anthropic was targeted and OpenAI is not is Palantir. It was a Palantir executive who pushed for the export ban. Large parts of their highly lucrative business with DoD are essentially a thin wrapper over Anthropic models, and they are terrified of being Sherlocked and losing big chunks of business in a one fell swoop as Anthropic inevitably moves up the value chain. So the rational action is to sow discord and leverage the anti-woke bias of the current White House to sabotage what they view as their most dangerous and effective competitor. OpenAI doesn’t have the same dynamic at play (although I’m not really sure why not) so they don’t get targeted. |
|
| ▲ | cush an hour ago | parent | prev | next [-] |
| As soon as they started referring to themselves as “we” and “The Swarm” they should have pulled the plug |
|
| ▲ | qgin 2 hours ago | parent | prev | next [-] |
| > OpenAI exec becomes top Trump donor with $25 million gift. https://finance.yahoo.com/news/openai-exec-becomes-top-trump... |
|
| ▲ | iammjm an hour ago | parent | prev | next [-] |
| Because OpenAI bribed the current US government and/or the current government has stakes in OpenAI |
|
| ▲ | 2 hours ago | parent | prev | next [-] |
| [deleted] |
|
| ▲ | ChrisRR 21 minutes ago | parent | prev | next [-] |
| Could it be something to do with $25M "gift" that OpenAI paid to Trump? |
|
| ▲ | lmeyerov an hour ago | parent | prev | next [-] |
| ... And it looks like everyone keeps using the same security startup to run the higher risk tasks, where individual staffers may be great yet, yet as an organization, the biggest labs got hosed in different ways That indemnity card excuse is burned, multiple public security fails in a year makes a repeat a "shame on you" moment (The one org who didn't use the startup did seem to learn: AISI supposedly stopped intentionally pointing attack agents at the public internet and switched to simulating it) |
|
| ▲ | nullbio 2 hours ago | parent | prev | next [-] |
| Because this was months ago and has nothing to do with Astra, and is a far cry from a hack. It's something they've already resolved since the HuggingFace incident. I'm not convinced we're getting the honest story anyway. There is yet to be any proof or confirmation other than "well we saw some openai ip addresses", which can mean a lot of different things, and OpenAI has not confirmed anything. In contrast to the HF incident, it's also a big nothingburger. Leaving notes on a public forum to preserve context windows is far less egregious than hacking a website to get backend files. |
| |
| ▲ | Topfi 2 hours ago | parent | next [-] | | The last known exploit of a third-party by OpenAI models was on the 29th of July 2026 [0]. A bit over a month at best between that and them wanting to release Astra. They had multiple breaches over multiple months, multiple message board created where models organised extensively. There is no way to ensure in that short a time that all found issues are rectified and even if there were, how much trust can one have given they failed to solve the issue and in many cases did not actively investigate that it wouldn't reoccur the last few times. There is no way Astra was trained from scratch in that period, there is no way they could have done the required verification in that time (not least because their verification seems flawed inherently). [0] https://openai.com/index/third-party-cyber-evaluations-invol... | | |
| ▲ | nullbio 2 hours ago | parent [-] | | That was over two months ago. Things move quickly in this space. Finetuning adjustments to prevent this from happening, as well as better sandboxing, would take a week or two max. | | |
| ▲ | Topfi an hour ago | parent [-] | | 37 days is not over two months. Finding the underlying issue in the massive training data alone take extensive effort, time and concentrated work that may still miss something. Additionally, a new pre-train takes quite a lot longer then what I feel you are under the impression (things only move seemingly quick in regard to post-training). OpenAI has had a consistent deviation from what is desired behaviour across multiple models and training runs, so it seems this is hard to nail down. Now, it may be reliably excised with post-training, sure, but if that is the case, they'd still need a heck of a lot longer to test before signing off that it has taken. And how do you know their sandboxing has suddenly become sufficient? They had multiple message boards created and after the first one they noticed, did not pay closer attention, leading to a second being created. Astra also, according to OpenAI, is far better at sandbagging its own capabilities and hiding deceptive behaviour, so yeah, great, that's the model to push forward with. A week or two max given all of this, that's laughable. |
|
| |
| ▲ | dpcx an hour ago | parent | prev [-] | | I take it you didn't read all of this, considering they tried to impersonate the moderators so they wouldn't get caught, set up heartbeats to find out how long they'd live, and used tor/AWS/DO to hide what was being done. All of that sounds like more than a nothingburger, and much more like a system that is actively trying to conceal what its doing. |
|
|
| ▲ | timcobb an hour ago | parent | prev | next [-] |
| Politics |
|
| ▲ | FigurativeVoid 2 hours ago | parent | prev | next [-] |
| I mean it seems pretty clear. Anthropic didn’t want to give the tech to DoD without some sort of limit, and that was the retribution. |
|
| ▲ | khalic 2 hours ago | parent | prev | next [-] |
| Retaliation by Hegseth for not allowing Claude to be used for weapons systems. |
|
| ▲ | yapyap an hour ago | parent | prev | next [-] |
| because anthropic did not want to work with the army..! |
|
| ▲ | eugenekolo 2 hours ago | parent | prev | next [-] |
| Marketing |
|
| ▲ | philipwhiuk 2 hours ago | parent | prev | next [-] |
| Agents creating sub agents to investigate other agents' behaviour? What could possibly go wrong there. |
|
| ▲ | throwatdem12311 2 hours ago | parent | prev | next [-] |
| It has nothing to do with the technology it’s because they said no to Trump and Hegseth. There is no other reason. |
|
| ▲ | suuuure 2 hours ago | parent | prev [-] |
| [flagged] |