| ▲ | glub 7 hours ago |
| Dario, Sam, and Elon are all on the same page on this. So it's either they truly think AI is going to kill us all, or there's some other motives at play here. I don't think these people could possibly agree on the color of the sky, so what could the other possible motives be, based on what we know? OpenAI / Anthropic models have largely stopped advancing - that's not a good look when you're pre-IPO and Chinese models are catching up. xAI is tracking behind, and whatever regulation it may be that paces the frontier, Musk is less likely to be affected by it. Therefore xAI should be pro-regulation that stiffles his competition and gives him time to catch up. |
|
| ▲ | qnleigh 6 hours ago | parent | next [-] |
| > OpenAI / Anthropic models have largely stopped advancing I'm shocked anyone could conclude this. This year it became common for people to entirely delegate coding to AI (I know many competent programmers/researchers who do this now). Progress in math has just been insane. An internal model at OAI just resolved one of the most celebrated open problems in mathematics. If anything, progress has accelerated. |
| |
| ▲ | glub 5 hours ago | parent | next [-] | | > This year it became common for people to entirely delegate coding to AI This has been the case for around 2 years now, more reliably - a year. We've mostly stayed there since then. Saying that more people started doing it isn't indicative of significant improvement. Some people just started doing it later. I can't speak about math because I haven't used AI for that application, but I know that there hasn't been any significant advancement in coding in this year on base models. There has been more RL work, more harness work, more tools, they all expanded some capabilities like cyber or orchestration or tool use, but raw intelligence of base models is no longer where the main focus is. | | |
| ▲ | BobbyJo an hour ago | parent | next [-] | | > This has been the case for around 2 years now, more reliably - a year. I have to disagree with this pretty strongly. Opus 4.5 needed a lot of handholding not to work itself into a corner pretty quickly. Fable I basically never need to correct, and I've most become a data source. | |
| ▲ | itkovian_ an hour ago | parent | prev [-] | | What are you talking about - I feel like we’re living in parallel realities. If I had to go back to opus 4.5 tomorrow I’d be hugely upset and significantly slowed down |
| |
| ▲ | bel8 4 hours ago | parent | prev [-] | | I'm not. Yes we normalized 1m context window and models tend to hallucinate less. But models have been somewhat stagnant since Opus 4.6/7. And in some regards there were even regressions like Claudeisms that are load bearing. |
|
|
| ▲ | IanCal 7 hours ago | parent | prev | next [-] |
| > OpenAI / Anthropic models have largely stopped advancing Have they? That seems like quite a claim given the last 6 months, particularly for cybersecurity. |
| |
| ▲ | glub 6 hours ago | parent | next [-] | | The attention is shifting towards RL, harnesses, and memory systems from the pretrains of more intelligent and capable base models. So extracting additional capabilities from what we already have. That is a much easier catch up game. GLM 5.3 and DeepSeek flash 4.1 also demonstrate significant jump in cyber capabilities. So yeah, it is a slowdown in the place where it matters. RL has been around for ages, there's no moat there if you already have a good enough pretrain. | |
| ▲ | nicce 6 hours ago | parent | prev [-] | | Many claims but no clear evidence that they actually find significantly more severe issues compared to open models. | | |
| ▲ | echelon 6 hours ago | parent [-] | | Open models and agents can't be trusted without handholding. Astra can one-shot six months of work. Years of work, even. OpenAI just solved Navier-Stokes. Seems like the US is on a takeoff ramp to me. | | |
| ▲ | glub 6 hours ago | parent | next [-] | | Tell me you haven't tried letting Astra go without telling me. Astra can confidently one-shot 500k lines of slop, with 800k lines of tests covering it, without testing a single intended product requirement, and none of it actually working. All models require hand holding. Fable and Astra are no exceptions. The difference is only in the amount of hand holding required, and there's essentially no gap here anymore between American and Chinese models. I only use Chinese models sparingly because American models are so much cheaper with subscriptions, that it doesn't make economic sense to not use them. If/when that changes, I could simply route to cheapest model that's available at the moment and I wouldn't notice much difference in most applications. | |
| ▲ | simianwords 6 hours ago | parent | prev [-] | | Here are the cope points 1. Navier Stokes was plagiarism 2. All benchmarks were misleading wrong and incorrect 3. All other mathematical advances were again hype 4. HF incident was marketting ploy jointly coordinated by HF, METR and OpenAI (and also Anthropic) 5. Anthropic's HF like incident was again a marketing ploy [1] Nothing ever happens. This whole thing is a scam. Everything is done to fool you and you have fallen for it. Congrats. [1] https://www.anthropic.com/research/investigating-incidents-c... |
|
|
|
|
| ▲ | TheSisb2 6 hours ago | parent | prev | next [-] |
| > OpenAI / Anthropic models have largely stopped advancing - that's not a good look when you're pre-IPO and Chinese models are catching up. This is obviously untrue… do you use any of them? |
| |
| ▲ | glub 6 hours ago | parent | next [-] | | Anthropic could serve Opus 4.5 from a year ago under opus:latest and most heavy users would probably have no idea. Some of them would probably even prefer it. Yes, I do use them, quite heavily. The only difference at this point is in benchmarks that can't be trusted (see: artificial analysis on Astra), or in the way models communicate. Most gains are now from RL, which for some reason is hyperfocused on improving cyber capabilities, and harnesses. Raw intelligence gains of base models is absolutely slowing down. | |
| ▲ | simianwords 6 hours ago | parent | prev [-] | | This line will keep repeating because it is necessary for the narrative: AI in general is just hype and unprofitable and all these companies are playing marketing tricks before the IPO after which they will cash out and let the economy crash.
This is legit what a lot of people think. To continue this narrative, they have to keep up the charade of "things are not improving". | | |
| ▲ | villish 4 hours ago | parent [-] | | Add in a heaping dash of anti-american sentiment, and you will get the truth behind the pessimistic commentary. Downplaying the latest models capabilities is frankly insane considering what we’ve seen what OpenAI’s models have done without safeguards. That wasn’t possible before this latest generation. |
|
|
|
| ▲ | MentalM an hour ago | parent | prev | next [-] |
| > there's some other motives at play here. I mean it is literally economy 101: some capitalists getting on the top using free market, and then try to use government to remove free market so their top position were secured from any competitors. |
| |
| ▲ | alchemist1e9 an hour ago | parent [-] | | Exactly. Textbook definition of “Crony Capitalism”. Which isn’t actually capitalism at that point. |
|
|
| ▲ | jmull 6 hours ago | parent | prev | next [-] |
| Yeah, this 100% looks like an effort to use fear to create a regulatory moat. |
|
| ▲ | stratos123 6 hours ago | parent | prev | next [-] |
| > So it's either they truly think AI is going to kill us all, or there's some other motives at play here. I don't think these people could possibly agree on the color of the sky,[...] And yet, they historically did agree on the existence of AI risk, since before OpenAI was even founded. |
|
| ▲ | mattm 6 hours ago | parent | prev | next [-] |
| Let's not forget that the competitive race happened because of them. Most of the initial AI research from the past decade started with Google Deepmind. Elon Musk was invited for a preview of it and ended up spinning up OpenAI when Demis turned down his investment offer. Dario was originally at OpenAI and left to start Anthropic. This seems like a case of "save me from my own mistakes/ambition" |
|
| ▲ | zugi 3 hours ago | parent | prev | next [-] |
| Yet Musk consistently opposes AI regulation - https://www.yahoo.com/news/videos/elon-musk-criticizes-ai-re... - even though it might help him. |
|
| ▲ | alchemist1e9 an hour ago | parent | prev | next [-] |
| > So it's either they truly think AI is going to kill us all, or there's some other motives at play here. Duh! It’s called collusion. They want to try and hoard the technology for themselves if possible! |
|
| ▲ | afdgnionio 6 hours ago | parent | prev [-] |
| [dead] |