| ▲ | porphyra 7 hours ago | ||||||||||||||||||||||
Why do they only host small models rather than the 2.4T version? Is the I/O and interconnect between the wafers bad due to the limited beachfront relative to the massive size of the chip? | |||||||||||||||||||||||
| ▲ | gardnr 7 hours ago | parent | next [-] | ||||||||||||||||||||||
They make a giant inference chip. Their inference service is basically just advertising for their core value prop: hardware. The CEO was on Gradient Dissent a couple years ago: https://www.youtube.com/watch?v=qNXebAQ6igs | |||||||||||||||||||||||
| ▲ | codexon 6 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
The wafer only has space for 44 gb of sram. If they offload ram they lose the speedup of having everything on 1 chip (the whole point of cerebras). | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | altertable 7 hours ago | parent | prev [-] | ||||||||||||||||||||||
Mostly economics I'm sure | |||||||||||||||||||||||