| ▲ | user43928 2 hours ago |
| Because DeepSeek is not "a month or two" behind as claimed in the article. These open models still did not beat February's Mythos / Fable 5. DeepSeek 4.1 Flash is behind GPT 5.6 Sol, and that one is left in the dust by the excellent Opus 5.5. Rumors say Anthropic is holding in reserve the big improvement, Fable 5.5, for the IPO. It's plausible that open models are 6 - 12 months behind, and there is no "good enough". As long as progress doesn't slow down, leading labs have nothing to fear. |
|
| ▲ | BobbyJo 2 hours ago | parent | next [-] |
| I was thinking about this earlier today and I came to the following question: If you had a model 10x as capable as the best model out today, but it cost 100x more, would there be a market, and, if so, how big? I think there would be a market and I think it would be large. So, I agree. |
| |
| ▲ | lifeisloving 31 minutes ago | parent | next [-] | | Many people would, and you'll find that they're building crappy webapps where you dont need SoTA. Like seriously who needs these frontier models? Unless you're doing some extermely difficult post-grad lvl research, you do not need a 100x PhD research assistant, especially not for whatever silly SaaS product most people are building. There's people at my job that get so much more done than everyone else using Fable/Opus/Astra. and all they use is the fastest cheapest models. I'd say the people who are using sota models for everything are doing it just because they prefer to be lazy. You simply do not need these frontier models, they outgrew most people's needs 6 months ago, but for some reason people still want to run a 700k rack of gpus full throttle to center a div for them. | |
| ▲ | ForHackernews an hour ago | parent | prev [-] | | Doing what? How many jobs involve solving Millennium Prize math challenges? 99% of everything is CRUD LoB apps. | | |
| ▲ | BobbyJo 37 minutes ago | parent [-] | | I am coding CRUD apps with a mix of astra, sol 6.1, fable and opus 5.5. A more capable model would still benefit me imo. Being able to follow high level guidance better, and being able to harness other models for each task would be a big improvement. | | |
| ▲ | lifeisloving 28 minutes ago | parent [-] | | Do you know how what you're doing, or do you find yourself working on things you dont understand and need the best model because it's the only way to push your own capabilities (because you're avoiding learning how to do the thing yourself)? Not asking to be mean, I just genuinely dont know why you'd need the frontier for basic applications. |
|
|
|
|
| ▲ | aleqs an hour ago | parent | prev | next [-] |
| that just sounds like openai/anthropic cope/propaganda, based on absolutely nothing objective lol even their harnesses are far surpassed by pi and opencode at this point also sick 'rumors' lmao, apparently marketing through rumors is in vogue these days |
| |
| ▲ | enraged_camel an hour ago | parent [-] | | >> that just sounds like openai/anthropic cope/propaganda, based on absolutely nothing objective lol Nah. There are benchmarks. They are free to look at. And they paint a very clear picture. | | |
|
|
| ▲ | ForHackernews an hour ago | parent | prev [-] |
| There absolutely is "good enough" and I agree with this author: DeepSeek 4.1 Flash is plenty good enough for all the things I would trust an AI to do at my job. |