| ▲ | nickysielicki 2 hours ago |
| The important take away here: the leapfrogging we’ve seen this year doesn’t seem to be a temporary thing. The famous theory of Dario Amodei was that AI was this winner-takes-all field where the first team to get a head start would never cede ground back. The term he liked to use was, “concentrating”. This is yet another datapoint that he was wrong about that. AI seems more distributed amongst neoclouds and traditional hyperscalers, FAANG and startups, GPUs and ASICs than it did this time a year ago. Nobody has a moat. |
|
| ▲ | LarsDu88 2 hours ago | parent | next [-] |
| Google has TPUs, a frontier model, a completely separate and lucrative revenue stream they can call on at will, and teams working on multiple different language modeling strategies simultaneously. Did I mention the vast and ominous data centers that already serve a significant fraction of the internet? If that ain't a moat, then what exactly is a moat? |
| |
| ▲ | throwitaway222 a minute ago | parent | next [-] | | Not to mention - they have the internet already indexed (they have a local copy), all books, and youtube. Beyond that, everyone "connects" to Gmail, but google HAS Gmail. | |
| ▲ | jobs_throwaway 2 hours ago | parent | prev | next [-] | | Then why have they been lagging behind OpenAI and Anthropic for most of the last few years, and only briefly been at the frontier? | | |
| ▲ | root_axis 29 minutes ago | parent | next [-] | | > Then why have they been lagging behind OpenAI and Anthropic Because they're not desperate. Slow and steady wins the race, at this rate all Google has to do is wait for OpenAI and Anthropic to exhaust themselves on aggressive training, then they can casually amble along right past them. | | |
| ▲ | pavlov 11 minutes ago | parent | next [-] | | This is basically what people were saying about Microsoft versus Google and other upstarts in the early 2000s. Slow and steady wins the race. How could they lose to something on the web, when Microsoft owns the web browser itself? Everything runs on Windows and IE. They can just relax and wait for competitors to exhaust themselves, then quickly build their own version. Isn’t that how Netscape lost. Etc. Now Google is the new Microsoft, just like Microsoft became the new IBM. | | |
| ▲ | brazukadev 5 minutes ago | parent | next [-] | | Yes Google is the new Microsoft, even Microsoft is still the old Microsoft but OpenAI is not the new Google. | | |
| ▲ | guelo a minute ago | parent [-] | | For some of us OpenAI is already the new Google. I prefer OpenAI even when I'm searching for web links. Google Search had already been deteriorating terribly before the rise of LLMs. |
| |
| ▲ | bamboozled 4 minutes ago | parent | prev [-] | | But Microsoft seems to be doing fine too ? |
| |
| ▲ | ivanmontillam 19 minutes ago | parent | prev [-] | | Also because they are Google. Google being Google: unreliable (they could kill a product anytime), too much of a platform risk (all products in a single place, get banned and lose the company), lack of support unless you really pay a big bill (into the 7 figures), among other… things. | | |
| ▲ | Traubenfuchs 9 minutes ago | parent [-] | | Incredible how the idea of Google being such bad choice has become so entrenched that people prefer the offerings of two companies that might not exist anymore in a few years over Google‘s similar offering. | | |
| ▲ | brazukadev 4 minutes ago | parent [-] | | > people prefer the offerings of two companies that might not exist anymore in a few years That's actually a feature. We don't need 2 new Googles. |
|
|
| |
| ▲ | gandreani a minute ago | parent | prev | next [-] | | Are you ready to call this race 5 years in? I'm waiting another 15 years and at least one or two new LLM architectures. | |
| ▲ | aleph_minus_one an hour ago | parent | prev | next [-] | | >
Then why have they been lagging behind OpenAI and Anthropic for most of the last few years, and only briefly been at the frontier? One possible explanation: because Google is a little bit more frugal and focuses on how to make providing AI models financially feasible - combined with some willingness to burn money so that they don't strongly fall behind on their AI models. On the other hand, OpenAI and Anthropic at least formerly concentrated on building and providing the best models that they could with concerns about financial feasibility taking a backseat. Just to be clear: I do have the impression that by now (likely because of pressure from investors) OpenAI and Anthropic take these financial concerns more seriously, but nevertheless Google's vs OpenAI's/Anthropic's "DNAs" concerning on what to focus on differ. | | |
| ▲ | monkeydust an hour ago | parent | next [-] | | And also Google is public listed company and the other two (for now) are private. I think that fact does have a strong bearing on how they operate. | |
| ▲ | ErrantX 5 minutes ago | parent | prev | next [-] | | I think this is the correct answer. Google never really has to outpace the competitors (other than to have some relevance) but they have a very large group of business customers using them for business process work in Gmail, Docs, etc. Clearly they will win when the models are close enough to frontier to be good enough, but are long-term cheap for buy. I.e. they will aim to make it a commodity. In theory MS has the same opportunity (plus they have GitHub so, you know, dev eco system too) but seem be blowing the strategy. Anthropic and OpenAI are having to race to the top on ability entirely to keep their name in the media and in front of us all (which costs: hence more recently trying to pivot away from model releases and more into controversy/danger). The main cost is in training and so this strategy is much much more expensive and this will play out either as a huge cost hike or a forced slow down in pace. I believe essentially Google is betting on that & I think it's probably the right strategy. | |
| ▲ | redanddead an hour ago | parent | prev | next [-] | | > Google is a little bit more frugal and focuses on how to make providing AI models financially feasible There’s no magic there. You get an account executive and a call with a systems architect to find out what you’re doing. Clouds gonna cloud, this is the reason they rolled deepmind into gcp and arguably the inverse is true, the labs are trying to become clouds | |
| ▲ | topspin an hour ago | parent | prev [-] | | > because Google is a little bit more frugal and focuses on how to make providing AI models financially feasible That feels right. It's not as if they've been missing out on great profits. |
| |
| ▲ | tfsh an hour ago | parent | prev | next [-] | | > Then why have they been lagging behind OpenAI and Anthropic for most of the last few years, and only briefly been at the frontier? Because it's not an existential battle for Google. If OAI or Anthropic disappear from the absolute frontier for ~8 months the news cycle and churn will diminish them to the second rate. Google is processing near 4 quadrillion tokens every month, that's - I'm sure - significantly more than OAI or Anthropic, because Google is interested more so in their flash models and getting these competitive, which they are. | | | |
| ▲ | merb an hour ago | parent | prev | next [-] | | From a business perspective a frontier model does not make much sense anymore if you are not a startup. Neither for Amazon, nor for Google. Their clouds need models that are fast and perform well in their agent frameworks nothing were a frontier model excels at. Most Google products even use flash lite underneath, so their frontier model is mostly used for distillation. | | |
| ▲ | aleph_minus_one an hour ago | parent [-] | | >
From a business perspective a frontier model does not make much sense anymore if you are not a startup. Neither for Amazon, nor for Google. Their clouds need models that are fast and perform well in their agent frameworks nothing w[h]ere a frontier model excels at. A good consideration; just one point from my side: as far as I am aware (but I may be wrong), Gemini is not known to perform well in an agentic framework. This is no contradiction to your other claims, quite the opposite: perhaps (or even likely) Google wants to avoid that their models become a commodity in some (agentic?) application where the middleman who actually writes this application gets a disproportionate of the money that the customer of the application pays for it. | | |
| ▲ | merb 5 minutes ago | parent | next [-] | | Well at the moment we do not use Gemini in an agentic framework. But as said the flash and flash lite families are basically their driving force in some of their applications in Google cloud, like document ai and its ocr capabilities are probably better when it comes to business documents than any other (at least in perf to cost to speed).
We also drive Gemini lite in our application where customer can use it to generate simple automation, like an agentic framework but way way smaller scale. And while it struggles in more context heavy operations it still is a beast when feeding it one or two pdf documents and asking questions about them and it’s hella fast. | |
| ▲ | nozzlegear an hour ago | parent | prev [-] | | > A good consideration; just one point from my side: as far as I am aware (but I may be wrong), Gemini is not known to perform well in an agentic framework. I used it for a month over the summer, right before they were going through the migration to antigravity. It was a fine workhorse IMO, no complaints from me. |
|
| |
| ▲ | krona an hour ago | parent | prev | next [-] | | Alphabet issued a very oversubscribed 100-year bond with 6.1% yield earlier this year to raise capital for datacenter expension. Meanwhile, Anthropic/OpenAI will struggle to survive the next 24 months on their current trajectory. | | | |
| ▲ | Yizahi 21 minutes ago | parent | prev | next [-] | | Lagging by what metric exactly? Is Toyota lagging behind McLaren? (company vs company) | |
| ▲ | spyckie2 30 minutes ago | parent | prev [-] | | Could it just be that Demis was checked out of the race and they lacked leadership? | | |
| ▲ | mattm 24 minutes ago | parent [-] | | Demis wants to focus on scientific endeavors. He was probably not the right person to focus on consumer apps. |
|
| |
| ▲ | IshKebab an hour ago | parent | prev | next [-] | | Don't forget the training data! Legal copies of all the books in the world, the entire web scraped, and all of YouTube. | |
| ▲ | bluecalm an hour ago | parent | prev [-] | | I don't think that other revenue stream is completely separate. It weighs on them as they need to think about tradeoffs. Classical search is going away sooner or later so they need to replace that with AI powered search. Data centers are important but a few others also has them: Amazon, Microsoft, Meta. SpaceX will likely be in/at the top I AI dedicated precessing power in 2027 as well. I don't see the moat. I see a company with a lot of other commitments that is not the best at delivering consumer facing products. They have some good cards but so do others. |
|
|
| ▲ | xnx 2 hours ago | parent | prev | next [-] |
| > Nobody has a moat. Custom hardware, data centers, huge cash reserves, deep/broad talent pool, and non-AI customer base are all huge advantages if not moats. Google, Microsoft, or Amazon are more likely to be the AI leaders than OpenAI or Anthropic. |
| |
| ▲ | dmix 6 minutes ago | parent | next [-] | | The technical crowd will always overvalue the recent technical advantage of software over the available customer base and business model math. | |
| ▲ | woah an hour ago | parent | prev | next [-] | | > Google, Microsoft, or Amazon are more likely to be the AI leaders than OpenAI or Anthropic. If not now, then when will these companies be AI leaders? Even Google, with its staggering advantages in cash, compute, real estate, training data, and having basically invented the field only manages to briefly claim a 1-2 week lead once or twice a year. | | |
| ▲ | koe123 an hour ago | parent [-] | | The financials for Anthropic and OpenAI are likely borderline suicidal, google and co are publicly traded. Moreover, all innovations downstream to them dont they? Why not just stay slightly behind, especially given many have stake in those other companies? | | |
| ▲ | lossyalgo 20 minutes ago | parent [-] | | Microsoft owns 51% of OpenAI, so they just have to wait for them go bankrupt then they come in and clean house. |
|
| |
| ▲ | bluGill an hour ago | parent | prev | next [-] | | There are many companies that have data centers. They are conceptually easy to build. An ASIC is difficult enough that if you make one someone will leapfrog you while you are still making it (at least so far), though once you have one your costs will be enough lower than the competition that you can perhaps undercut them. | | |
| ▲ | xnx an hour ago | parent [-] | | True, but have other hyperscalers caught up to Google's AI data centers?: fully liquid cooled, torus networking(?), 100,000+ TPUs interconnected, etc. Google is already on gen 8 of its TPUs and is certainly already working on the next version or two. |
| |
| ▲ | IX-103 2 hours ago | parent | prev | next [-] | | If Moore's law continues, then in less than 10 years today's state of the art model will be able to run on a cell phone. How much smarter do we actually need AI to be? Would it still require datacenters and custom hardware? | | |
| ▲ | xnx an hour ago | parent | next [-] | | Moore's law stalled ~2015. Unfortunately, no way current models will run on the <100W thermal budget of a cell phone. Printing the weights directly into a chip would help efficiency a lot, but not enough. | | |
| ▲ | CuriouslyC 43 minutes ago | parent | next [-] | | In all likelihood in a few years we'll get ~200-400bA~4-6 MoE models that are on chip, and they'll be better than the current frontier. | |
| ▲ | spacebanana7 22 minutes ago | parent | prev [-] | | Is it conceivable that in 10 years time we’ll have 7B models that have the same level performance as modern frontier ones? |
| |
| ▲ | LightBug1 an hour ago | parent | prev [-] | | They probably said the same thing about social media back in the day. I'm sure the thinking out there, and hence investment, is all about how to tether the user to the most addictive, network-effected, incredibly deep, server-side, moat-able version of AI possible. |
| |
| ▲ | junehwi 2 hours ago | parent | prev [-] | | [dead] |
|
|
| ▲ | altruios 2 hours ago | parent | prev | next [-] |
| > Nobody has a moat except nvidia For now, for cloud training. but for consumers, nvidia vs amd reasonably close - the moat there is thin and shrinking. I suspect AMD will surprise us. nvidia has no motes in china, which may be a new source of (gpu) chip design. Huawei's Ascend 910C is about a generation behind... again: for now. point is: moats dry up. I see nvidia's shrinking as a real possibility. |
| |
| ▲ | culi 2 hours ago | parent [-] | | China will always be generations behind until they crack domestic EUV | | |
|
|
| ▲ | pvab3 an hour ago | parent | prev | next [-] |
| The whole winner-take-all idea seems entirely based around Singularity/Rationalism and would require massive advances that we probably aren't close to at all. |
| |
| ▲ | nater5000 39 minutes ago | parent [-] | | Yeah, it kind of seems like we haven't gotten to the "head start" he's referring to yet. | | |
| ▲ | laybak 16 minutes ago | parent [-] | | yeah this would be one interpretation where "winner take all" could still be right. though with the recent accelerating release cadence, doesn't it seem like the head start / lead has been shrinking over time? |
|
|
|
| ▲ | FinnKuhn an hour ago | parent | prev | next [-] |
| My personal theory is (assuming there really is no moat) whoever starts the latest with developing AI models might actually win as they should be able to develop a competitive product with significant less resources and initial investment resulting in a higher ROI. AI might even become a commodity. |
| |
| ▲ | koe123 an hour ago | parent | next [-] | | This is my secret hope for Europe! | |
| ▲ | pkfz 38 minutes ago | parent | prev | next [-] | | Given that the infrastructure won't be a moat and will become a commodity. | | |
| ▲ | FinnKuhn 25 minutes ago | parent [-] | | Based on historical developments the cost of compute will go down again eventually, decreasing the cost of training AI models of the same quality as today even further. That part is what I would be the most certain about. |
| |
| ▲ | dgellow an hour ago | parent | prev [-] | | That’s what we see in China |
|
|
| ▲ | vb-8448 2 hours ago | parent | prev | next [-] |
| It's even worse, we are crossing over into the realm of religion. The article against GML 5.3 is the equivalent of a Papal excommunication. |
| |
| ▲ | zone411 an hour ago | parent | next [-] | | The article presented facts and data. If that's a problem for you, that sounds more faith-based than whatever Anthropic is doing. | |
| ▲ | verdverm 2 hours ago | parent | prev [-] | | which article? have not seen this one --- maybe it's this Anthropic post on GLM? https://www.anthropic.com/research/glm-5-3-and-the-spread-of... > Governments should conduct safety testing on sufficiently capable AI models, including successors to GLM-5.3. Without high-quality evaluations from independent sources, the impact of these capabilities might not become fully clear to model developers until it is too late. As AI developers across the world build increasingly capable open-weight models, we hope they work to appropriately safeguard these capabilities and prevent misuse. I for one do not think my government is up to the task of designing or implementing such a system | | |
| ▲ | Rzor 2 hours ago | parent | next [-] | | https://www.anthropic.com/research/glm-5-3-and-the-spread-of... | |
| ▲ | Iolaum 2 hours ago | parent | prev | next [-] | | It doesn't need to. It can use your cash to pay the people who are. Those in power like it more that way. | |
| ▲ | esafak 2 hours ago | parent | prev | next [-] | | You say that because the 'most' existing models have done is hack governments and companies. Can't you think of worse things a model could do; accidentally or by instruction? | | |
| ▲ | verdverm 2 hours ago | parent [-] | | help people with suicide and school shootings like ChatGPT already has OpenAi is alledged to have been monitoring these internally and not contacting authorities. Lawsuits have been filed, I see gross negligence without the gory details I have for more concerns around human-chatbot maladies than I do around the cyber security stuff. For example, why hack grandma when you can get her to do something willingly through impersonation. How do we prove authenticity in a post truth world? |
| |
| ▲ | 2 hours ago | parent | prev [-] | | [deleted] |
|
|
|
| ▲ | kushalpandya an hour ago | parent | prev | next [-] |
| Even Google itself stated (internally at least) that nobody has a moat https://newsletter.semianalysis.com/p/google-we-have-no-moat... |
| |
| ▲ | bitpush an hour ago | parent [-] | | Wasnt it just some dude writing a doc? That's hardly a Google (The Company)'s position. | | |
| ▲ | dgellow an hour ago | parent [-] | | If anyone other than NVIDIA has a moat, they for sure never talk about it |
|
|
|
| ▲ | hirako2000 2 hours ago | parent | prev | next [-] |
| And before him, Altman was explaining very calmly that no company could ever compete with OpenAI. |
|
| ▲ | aleph_minus_one 2 hours ago | parent | prev | next [-] |
| > The famous theory of Dario Amodei was that AI was this winner-takes-all field where the first team to get a head start would never cede ground back. This is the kind of story that ones tells to investors to justify the huge amount of cash burn. :-) |
| |
| ▲ | mapontosevenths 2 hours ago | parent | next [-] | | I'm not sure it's wrong. This all feels a bit dotcommy to me. I think many/most of the players will crash and burn, and the ones that are left will divide the world. | | |
| ▲ | CoolestBeans an hour ago | parent | next [-] | | The problem is twofold. One, even a monopoly AI provider wouldn't have pricing power against its suppliers. Its suppliers are energy, semiconductors, and real estate. Semiconductors maybe they could get some leverage on but energy and real estate have plenty of other buyers. Two, there's still no evidence of a runaway scenario (ie a small lead turns into a big lead over time) and there's still no evidence that there's some resource that you can deny everyone else that they can't build your product also. You can't hoard energy, compute, memory, data, human talent, or customers. The net effect is that the most likely scenario is if one big lab fails, they will likely all fail. Their revenues are all correlated. To go to your dotcom comparison, the winner will be the ones picking through the assets that were written down by orders of magnitude and trying new products with the technology until one sticks to the wall. But I don't know if a dramatic crash is guaranteed either. | | |
| ▲ | aleph_minus_one an hour ago | parent | next [-] | | >
The problem is twofold. One, even a monopoly AI provider wouldn't have pricing power against its suppliers. Its suppliers are energy, semiconductors, and real estate. Semiconductors maybe they could get some leverage on but energy and real estate have plenty of other buyers. Concerning the leverage on energy and real estate: don't forget that the AI companies have quite a lot of choice where to build their data centers. So AI companies have lots of opportunities to play several parties off against each other (in particular also for real estate and energy). | |
| ▲ | 3d2 an hour ago | parent | prev [-] | | "but it can stay there so long as the balance sheet doesn't deteriorate." Uhm, what? LOL. People dont value firms based on balance sheets fella. Have you taken a basic valuation class? Tesla is a nice stock for traders - they like the volatility. Nobody holds Tesla as stock for investing. If you were to truly value it on an intrinsic value basis you'd have to bring in failure risk. |
| |
| ▲ | jaggederest 2 hours ago | parent | prev | next [-] | | I suspect this is going to end up like most services provided e.g. cloud stuff, balkanized between a couple major players and an assortment of DIY or less popular options if you don't like those ecosystems, plus some UX/DX focused wrappers that use the big players under the hood. I think that would be a pretty satisfactory outcome compared to one hypercompany consuming trillions of dollars of the world economy. | |
| ▲ | skybrian an hour ago | parent | prev | next [-] | | “Divide the world” sounds ominous. Here’s another scenario to consider: Internet access is not really unlimited, but for many people with fiber at home, it effectively is and we pay a flat rate. Perhaps by the end of next year, most programmers will stop thinking about metered access for AI? For many people, the cheaper models (about as good as today’s frontier models) will be good enough. Which might sound good, but the downside is that it will also be easier to build an AI botnet without the users paying for it noticing. Particularly when people are running AI inference on their own hardware. | | |
| ▲ | pixl97 39 minutes ago | parent [-] | | Hardware is still insanely hard to get a hold of, and the stuff that's being built doesn't really work for home use. Maybe if it crashes Nvidia will adjust the hardware flow. My guess is even if the AI market busts there is still a massive demand for hardware as models are solving all kind of problems now. But ya, lots of hardware everywhere not managed well is how you get sovereign AI. |
| |
| ▲ | pianopatrick an hour ago | parent | prev | next [-] | | Or, like airlines, the ones that are left will have great technology but be not so great from a business and financial perspective. To me AI seems like a commodity service. | | |
| ▲ | pvab3 an hour ago | parent [-] | | Like airlines but starting off with hundreds of billions of dollars of obligations and debt |
| |
| ▲ | 3d2 an hour ago | parent | prev | next [-] | | THe problem with analogies is that they are imperfect. I would argue those who already rule the world, will continue to do so. What happens to OAI and Anthropic? No idea, probs go bust. Google just has to offer a half-decent offering in the long run and have a cost-advantage and it'll eventually knock OAI and Anthropic out as firms figure out what combination of models they want to be best for their economics and generating returns. Enterprises trust google over OAI and Anthropic. A clear signal of this was the Apple deal. Dont forget those sweet returns fellas! CEO's are hired to make the owners wealthier. That is not gone. | |
| ▲ | 2 hours ago | parent | prev [-] | | [deleted] |
| |
| ▲ | ehsankia 2 hours ago | parent | prev [-] | | I guess if one of them hits singularity, it could in theory just wipe out all the rest, seeing how they keep escaping and hacking into other systems :) | | |
| ▲ | aleph_minus_one 2 hours ago | parent [-] | | >
I guess if one of them hits singularity, it could in theory just wipe out all the rest The story that some AI company might reach singularity and then "everything will be different" is another science-fiction story that executives of AI companies love to tell to justify the staggering amount of necessary investments and cash burn. :-) | | |
| ▲ | koe123 an hour ago | parent [-] | | I find it quite unique how many people buy into this. Its the worlds most blatant conflict of interest, I dont even know why Sam and Dario bother doing interviews | | |
|
|
|
|
| ▲ | grababner 21 minutes ago | parent | prev | next [-] |
| It's hard to make predictions, especially about the future |
|
| ▲ | SwellJoe 2 hours ago | parent | prev | next [-] |
| I think some in the AI industry drank their own Kool-Aid. They believed that if they had the best model and the most compute, they could tell the model, "Make a better model." And it would, and the next one could make its replacement, and so on. So far, that's not exactly how it's played out. Humans are still necessary for the leaps in capability or efficiency. A model can grind on a problem to eke out the most performance, and models can synthesize data and iterate on various techniques to find the optimal combination. But, seems like humans still have to provide the real thinking, and the talent and drive for doing that is not concentrated in one company or city or even one country. And, (surprisingly) a lot of the people involved are in it for advancing the field more than making another billion dollars, so they're publishing their research. So, yeah, the moat isn't deep. Even the compute moat, that OpenAI, Musk, and a bunch of other also-rans (like Oracle) bet the farm on, isn't really panning out. The Chinese makers just spent their effort on making models vastly more efficient, since they couldn't do anything about having an order of magnitude less compute available. |
| |
| ▲ | scottyah 2 hours ago | parent | next [-] | | But that's the whole point of the singularity. Right now the models use a lot of human effort and ingenuity to improve the models, but about a year ago it was 100% human. We'll see in another year, but if this pace continues I doubt there will be more than a handful of people who can contribute more than the models. | |
| ▲ | TeMPOraL 2 hours ago | parent | prev [-] | | > I think some in the AI industry drank their own Kool-Aid. They believed that if they had the best model and the most compute, they could tell the model, "Make a better model." And it would, and the next one could make its replacement, and so on. They're not there yet. Once they get there, that's literally the definition of Singularity. But they are getting closer. Recursive Self-Improvement used to be a phrase people mocked LessWrong crowd for using and worrying about, now it's something both OpenAI and Anthropic already publicly admitted not only to pursue, but to already be benefiting from. | | |
| ▲ | SwellJoe an hour ago | parent | next [-] | | Sure, it's happening...but, is IT happening? By that, I mean, we can see that the models are able to iterate at a pace and scale that humans can't match, and that provides gains in model performance and efficiency. But, humans are still needed in the loop, and not just because it's necessary for safety/alignment reasons. I don't think any significant discovery has been made by models on their own, and I don't know that LLMs will ever have the capacity to invent. They can synthesize from known data amazingly well, and since they know everything "known data" is extremely broad. But, the leaps, so far, have all come from humans. So far, I don't think the models are capable of running away on their own. Of course, it would be playing with fire to not at least consider the risks of such a runaway scenario and build in safeguards against it. But, there is no model that can build a better model on its own, thus far, to the best of my knowledge (which is far more limited than the models, so maybe I should ask them). | | |
| ▲ | TeMPOraL 33 minutes ago | parent [-] | | Recursive Self-Improvement isn't instant, it starts slow and accelerates. It starts with what they already claim to be doing - increasingly relying on existing models in non-trivial work related to training, evaluating and optimizing the next, more capable generation of models. As long as the proportion of work keeps shifting towards agents doing more and more of it, and humans less and less, that's RSI at play. It may be that it turns out LLMs lack some fundamental level of judgement and it plateaus, but frankly I find this notion absurd; LLMs already show better judgement than most people. The alternative is, at some point LLMs will show the ability to futz their way into improvement of the next generation of models even without humans in the loop - even if much less efficient at first, if generation N+1 is more capable than generation N, it'll either take off or burn out. |
| |
| ▲ | pvab3 an hour ago | parent | prev | next [-] | | Those people have a lot of overlap with the LessWrong crowd. They do not have RSI now and probably never will | | |
| ▲ | TeMPOraL 38 minutes ago | parent [-] | | They absolutely do, unless you believe they are lying about the fact they're using current generation models extensively to develop the next generation of their models. | | |
| ▲ | Yizahi 14 minutes ago | parent [-] | | The letter S stands for "Self" and word "using" is for sure not a superset of the word "self". Basically LLMs are assisting someone who does improving of said LLM, while RSI is a carpal... ahem, RSI is "self" improvement, meaning no intemediary in a human form. PS: it's also not recursive but iterative improvement, even if it ever happens. |
|
| |
| ▲ | thmoonbus an hour ago | parent | prev | next [-] | | The companies whose insane valuation is based on accomplishing thing X say they’re getting closer to accomplishing thing X? At least they’re led by trustworthy and honest people or we’d need to take their claims with some dose of skepticism. | |
| ▲ | woah an hour ago | parent | prev [-] | | Is it recursive self improvement if Claude Code writes your pytorch for you? | | |
| ▲ | TeMPOraL 39 minutes ago | parent [-] | | If the point of that pytorch is to improve the next generation of Claude, then yes, absolutely. |
|
|
|
|
| ▲ | Heidaradar 2 hours ago | parent | prev | next [-] |
| I just find this unlikely personally, think about the great research that's happening in the open source world, I'm sure inside anthropic + openai they've also made a bunch of discoveries and improvements (and I'd guess way more due to them attracting the best talent + the better internal models they have) |
|
| ▲ | ddp26 an hour ago | parent | prev | next [-] |
| Isn't the important takeaway here that Gemini 4 is not released and has no planned release date? This is marketing from Google, not a competitive offering |
|
| ▲ | qgin 2 hours ago | parent | prev | next [-] |
| Whoever gets to RSI first “wins” but also maybe ends life on earth. The incentives have never been worse. |
| |
| ▲ | pianopatrick an hour ago | parent | next [-] | | Maybe. Or maybe having the best AI model on the planet becomes like having the best super computer on the planet. Useful for some niche stuff, but not too useful in terms of people's daily lives or what is used in business. | | |
| ▲ | pixl97 27 minutes ago | parent [-] | | Why does your outcome make any sense? Already businesses that have more compute and access to data seem to eat the world around them. If, and ya its and if, we can make something that self learns into RSI it's not looking like any business that came before this. |
| |
| ▲ | koe123 an hour ago | parent | prev [-] | | RSI being science fiction so far. Whoever builds the deathstar wins! |
|
|
| ▲ | culi 2 hours ago | parent | prev | next [-] |
| I'm not necessarily defending this obvious marketing speak but maybe the "starting point" was wider than assumed. So far, nobody has caught up to US and Chinese labs for example despite lots of funding in Europe. This is also despite abundant in-depth research papers being published alongside open source code and weights by some Chinese labs |
| |
| ▲ | funnym0nk3y 2 hours ago | parent [-] | | There is not really much funding in Europe. At least not for start-ups. There is simply not enough compute in Europe. | | |
|
|
| ▲ | RachelF 2 hours ago | parent | prev | next [-] |
| The US companies still have trillion dollar valuations like there is a monopoly. There just isn't one. They are all within a few percent of each other on the benchmarks. The slightly lower Chinese open models are good enough for almost everything, too, and much cheaper. Like with humans there is plenty of employment for people with below genius level IQ's. |
| |
| ▲ | jppittma an hour ago | parent | next [-] | | I feel like the frontier labs are going to serve fast/lower intelligence models at a better per token cost than the open chinese models. You're telling me that in the long run, you're going to self-host your own ai infra for cheaper than google can serve it to you? I don't really buy it. I think the dedicated AI data centers are going to serve AI at a lower marginal cost than random businesses self-hosting, and then it's a question of how much of that margin they can capture. | | |
| ▲ | usef- 9 minutes ago | parent | next [-] | | Yes, it may come down to how well they get efficiencies from scale. People always compare the inflated API prices, but subscription prices of American models are competitive for the intelligence. You get >20x the subscription cost in tokens. | |
| ▲ | msy an hour ago | parent | prev | next [-] | | Agree entirely but that's the point, if it's a margin knife-fight with marginal product differentiation/pricing power nobody is going to be making bank. | |
| ▲ | pvab3 an hour ago | parent | prev [-] | | If they have to recoup training costs then they don't have much choice |
| |
| ▲ | fumar 2 hours ago | parent | prev | next [-] | | Is there a dividing line between good enough and best in class capabilities? It's blurry from where I stand. Will model makers cede ground or is there a market making moment up for grabs (singularity)? | |
| ▲ | handfuloflight 2 hours ago | parent | prev [-] | | > Like with humans there is plenty of employment for people with below genius level IQ's. Not if the genius level IQs take the market share. |
|
|
| ▲ | nylonstrung 2 hours ago | parent | prev | next [-] |
| I think that scenario only naively made sense if technical knowledge was entirely proprietary and talent was guarded with severe non-competes and NDAs And Chinese labs openly publishing so much of their methodology destroyed any hope, which was inevitable |
| |
| ▲ | torginus 2 hours ago | parent [-] | | > And Chinese labs openly publishing so much of their methodology destroyed any hope, which was inevitable I think the secrecy doesn't make sense. People swap jobs between labs so I'd say the big players can' really keep secrets for long, and any secret sauce advantage gets incorporated by competitors in a major product cycle at most. |
|
|
| ▲ | arizen 2 hours ago | parent | prev | next [-] |
| Seems like learning rate velocty may be the ultimate moat |
|
| ▲ | zem 2 hours ago | parent | prev | next [-] |
| I have never understood the whole "this is a winner take all game" mentality - the sheer size of the pie is so great that from a purely rational standpoint companies should just be trying to productively get a slice of it and be profitable. winner-take-all is just greed/capitalism run amok, where it is not enough to be profitable, you have to own the entire market (and presumably extract rents) |
| |
| ▲ | dgellow 41 minutes ago | parent [-] | | It’s like the supposed first mover advantage OpenAI believed they had. In practice it’s almost always more like a first mover massive tax, and companies coming afterwards benefit from your discovery of a market, publicity, and everything else that has already been validated |
|
|
| ▲ | Aboutplants 2 hours ago | parent | prev | next [-] |
| I feel his theory depends on the premise that access to pure compute would the be the determining factor of success. Not the case |
|
| ▲ | sixo an hour ago | parent | prev | next [-] |
| Nobody has a moat so long as employees can move between companies |
|
| ▲ | tripleee 2 hours ago | parent | prev | next [-] |
| Was Dario's company winning at that point in time by any chance? |
|
| ▲ | esafak 2 hours ago | parent | prev | next [-] |
| The present leapfrogging is not a contraindication because companies are not necessarily releasing their best models; we know they have smarter internal models. Furthermore, humans are still involved in model creation. Human involvement is expected to decrease over time, and when model iteration is completely automated, progress will happen at the machine's pace, leading to runaway intelligence, barring any ceilings. |
|
| ▲ | SecretDreams 2 hours ago | parent | prev [-] |
| AI is a commodity. One that is showing to be more readily commoditized than most has anticipated. As of now, the only moats are the financing for the hardware to run it and the hardware vendors themselves - with the latter largely not yet a commodity because of ecosystem lock and a limited capacity of the most advanced fabs in the world. |