Remix.run Logo
▲ nbardy 2 hours ago

I think it's weirdly just a choice of deciding to cut releases.

We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

▲233mhz 43 minutes ago | parent | next [-]

> We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

I mean, isn't it almost a guarantee that what we get is a gimped version of what they use internally? They probably already serve themselves next gen level models at 1k+ tps from cerebras machines hosted on perm while we get quantized astra/opus at 50tps on a good day

▲iLoveOncall 2 hours ago | parent | prev | next [-]

> We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

You're just believing their own bullshit. There's no indication that this is true except from claims from people working at OpenAI.

If they really had a much more powerful model, it would make absolutely no sense to sit on it.

▲meowface an hour ago | parent | next [-]

With all due respect, you do not have a single clue what you're talking about.

▲owebmaster an hour ago | parent [-]

With the same respect, you don't either. Simping for openai don't make you part of their in group

▲meowface an hour ago | parent [-]

I definitely do not know what I'm talking about, but that individual doesn't know what they're talking about even more than I don't know what I'm talking about.

▲233mhz 42 minutes ago | parent | prev | next [-]

> If they really had a much more powerful model, it would make absolutely no sense to sit on it.

Makes total sense if they don't have the compute and can't serve it in an economically viable way. Also it lets you build things no one else in the world can build as fast as you until it's released

▲iLoveOncall 9 minutes ago | parent [-]

> Also it lets you build things no one else in the world can build as fast as you until it's released

"We can generate slop faster than anyone in the world" :evil_emoji:

▲adamzenith an hour ago | parent | prev | next [-]

You don't think having a more intelligent model they can use internally that others can't is an advantage?

▲owebmaster an hour ago | parent | next [-]

If that's the case, do Anthropic have an even better one helping them? The Chinese labs too? Where this OpenAI "advantage" is taking them?

▲iLoveOncall an hour ago | parent | prev [-]

No? The top of human engineers are much better than any model would be, so AI models really aren't a big advantage when you're trying to develop anything that is SOTA.

▲ceejayoz an hour ago | parent [-]

Not every problem is best addressed by a top engineer.

Plenty of the tasks that keep a company running can benefit from good-enough (and better than the competition).

▲iLoveOncall an hour ago | parent [-]

Yes, and none of those tasks require even the current SOTA models.

▲simonw an hour ago | parent | prev | next [-]

It makes sense for them to sit on it until they've finished testing it. More powerful but also more likely to delete all your email by mistake = you shouldn't release it yet.

▲howdareme9 an hour ago | parent | prev [-]

its not done training, why would they release a model that hasn't finished training?

besides, we know anthropic are sitting on models too

▲iLoveOncall an hour ago | parent [-]

> its not done training, why would they release a model that hasn't finished training?

Because clearly they have no problem with releasing newer versions of models even just a week apart.

> besides, we know anthropic are sitting on models too

It is from your crystal ball or from other bullshit you heard from Anthropic employees on Twitter?

We all know Anthropic had Mythos and Fable, and they turned out to be completely normal models, entirely in line with the capability of their predecessors.

All they do is lie, and you're believing their lies.

▲meowface an hour ago | parent | next [-]

The poster is not claiming Bel is a secret AGI. Just that it exists and only exists internally at the moment.

It's rumored to be over 10T parameters. When released it'll probably be very good at certain tasks, albeit slow and expensive and not necessarily "wiser". You don't have to make this a binary.

Also, Mythos was in fact a significant step-up in several ways. It fits the trend line, but only because the trend line for LLMs is quite steep. Plus Astra is still in many ways less intelligent than Fable/Mythos despite being released much later.

▲azan_ an hour ago | parent | prev [-]

> its not done training, why would they release a model that hasn't finished training?

> Because clearly they have no problem with releasing newer versions of models even just a week apart.

You can see how it is pure non-sequitur, right?

> We all know Anthropic had Mythos and Fable, and they turned out to be completely normal models, entirely in line with the capability of their predecessors.

Fable was absolutely not in line with capabilities of other models when released. For cybersec work it was much, much better.

▲mFixman 2 hours ago | parent | prev [-]

Any strong enough model with weak enough safeguards can cause an AI Chernobyl event that will make people and governments against AI development and deployment, just like Chernobyl did for nuclear energy.

▲ChromeUltron an hour ago | parent [-]

tell me "I drank the kool aid" without telling me you drank the kool aid.

▲mFixman an hour ago | parent [-]

The US government and most large companies drank the kool aid, and they will be the ones blaming Big AI if things go very wrong.

▲esseph 29 minutes ago | parent [-]

I think there's going to be a brutal backlash against the entire technology sector.

Gov and Corp will throw their hands up and explain how it's not their fault.