| ▲ | simonw 5 hours ago | |
> We match DeepSeek V4 Pro Base using ~50x fewer FLOPs – that’s around half of GPT3’s pretraining compute, or ~$0.5M on GB200. If this holds up that's a really big deal. | ||
| ▲ | wayfwdmachine 5 hours ago | parent [-] | |
Huge if true. As it were. | ||