Remix.run Logo
ericd 3 hours ago

>You couldn’t buy one of these if you wanted to right now.

You can: https://www.exxactcorp.com/Exxact-TS4-149591758-E149591758 . You can get thousands of tps of GLM 5.3 output out of this thing, which grades around Opus 4.8. Payoff is around 1 year vs. spot prices on these GPUs, including power.

Aurornis 2 hours ago | parent | next [-]

> You can: https://www.exxactcorp.com/Exxact-TS4-149591758-E149591758 .

No, you can get a quote for possibly being allocated one in the distant future.

The backlog for these is huge. You cannot buy one any time soon.

ericd 2 hours ago | parent [-]

Ah gotcha. Have you tried to order something like this in the past?

arjie 2 hours ago | parent | prev | next [-]

I have quoted large nodes from this supplier and have lots of^W^W GPUs from them for personal use. Current lead time is more than 30 months.

They're a good provider but you have to be a big shot buying NVL72s before you're getting anything within your payback period.

ericd an hour ago | parent [-]

Ah thanks for the solid info, too bad. I'd seen them come up as a pretty good price for 6000 RTX's in the past, which seem generally pretty available, good source for those?

arjie 41 minutes ago | parent [-]

Yeah, they're good source. But the price for those GPUs is 5 figs even with the nvidia startup program nowadays. Also, I went back and looked. Most of my GPUs are actually from Central Computers who were great, but Exxact is real too. So "lots of" was inaccurate.

Also, the lead time I quoted was for individual 8x nodes.

ericd 32 minutes ago | parent [-]

Ah yeah, one of mine is from Central. And yeah, crazy how much they've gone up. But I can see why, they scream.

CamperBob2 2 hours ago | parent | prev [-]

I can't tell from the ad -- it says "supports" 8x MI350X GPUs, but does that mean "includes" 8x MI350X GPUs? For $300K I'd certainly hope so, but I'm assuming not.

A system with 4x RTX 6000s costs about $60K these days, and can (as you note) trade blows with Opus 4.8 if not Fable. In fact, it'll give you a better pelican than Fable 5.1, and in less time.

ericd 2 hours ago | parent | next [-]

Ha fair, I'd definitely confirm with a salesperson before wiring them $300k. But most of the signs on the configurator seem to point to it including the GPUs? Not going to make 30k BTUs/hr of heat without the 8kw of GPUs.

Aurornis 2 hours ago | parent | prev [-]

> trade blows with Opus 4.8 if not Fable.

Okay I love the open models, but the hype is getting ridiculous. The models you can run on 4 X RTX6000 are not Fable level.

ux266478 an hour ago | parent | next [-]

Baseline yeah. But part of the reason you run open models is how much nicer fine tuning them is. Granted, you probably don't want to try and make LoRAs on a 4x RTX6000 setup, but you could if you really wanted to and there are other ways to modify models. And yes, if you're good at it, you can turn a piddly mid-range model that's only good at benchmarks into a heavyweight clanker (for a specific domain).

CamperBob2 2 hours ago | parent | prev [-]

Well, they are if you're into animating pelicans. :-P But yes, in the general case Opus is a better match.

And Opus is no slouch. I'm satisfied that GLM 5.3 is just as strong as Opus. Z.AI has promised/bragged that they will be at Fable 5.0 level by the end of the year or early next year, and I don't see any reason to doubt them.