Remix.run Logo
nh43215rgb 3 hours ago

  > Built on a sparse Mixture-of-Experts architecture, Step 5 Preview has 600B total parameters, with 27B active per token, and supports a 1M-token context window and vision input.

  > Step 5 Preview scores 44 on the Artificial Analysis Intelligence Index.

  > The model will be released with open weights on October 15.
I guess being Chinese company they decided to skip version 4, while also giving impression to be on the similar iteration with leading companies (claude opus 5). I wonder if other Chinese labs like Kimi/Moonshot will follow suit.
JohnsonZou an hour ago | parent | next [-]

Another possible reason is that the number 4 is considered unlucky in traditional Chinese culture.

Bolwin 2 hours ago | parent | prev [-]

Moonshot has already teased K3.1 so not likely

dannyw 2 hours ago | parent [-]

K3.1 would likely be a deeper/longer post-train from K3, so that’d make sense.

It’s all marketing anyways, but that’s at least how a lot of labs have been naming things (sometimes).