| ▲ | Luker88 7 hours ago | |
Mistral Large 4: 1050B, 49 Active GLM-5.3: 753B, 40 Active I was hoping for something that hinted at smaller models too, but I guess not. Any competition is still good, especially now that the USA AI labs are starting to do regulatory capture. | ||
| ▲ | rahen 5 hours ago | parent | next [-] | |
Give them time. ML 4.0 was just pretrained. Mistral will certainly use it as the base for distillation and RL for smaller, better, more efficient iterations, just as the competition does. | ||
| ▲ | gandreani 6 hours ago | parent | prev | next [-] | |
It's competitive! Good enough to show competence, and instill confidence in the team/company. Later releases can be more efficient. I think it's a great release with that framing. | ||
| ▲ | nolok 3 hours ago | parent | prev | next [-] | |
It's Mistral Large, they usually publish Medium and Small later | ||
| ▲ | baq 6 hours ago | parent | prev [-] | |
just keep RL frying it should get better... | ||