| ▲ | manlymuppet 5 hours ago | ||||||||||||||||||||||
Man, a lot of this discussion sounds like people cheering for the last kid crossing the finish line. Surely we want competition and Europe involved in that, but at this point I have grown used to either American labs smashing the frontier remarkably fast, or Chinese labs getting way, way closer than you would expect them to. Mistral’s progress, regrettably, feels much slower. This model doesn’t knock anybody’s socks off. The model is (and I hate to be this harsh) mediocre, and this mediocrity has also arrived months late. This is a pretty grim prognosis for European AI. | |||||||||||||||||||||||
| ▲ | troyvit 3 hours ago | parent | next [-] | ||||||||||||||||||||||
> Man, a lot of this discussion sounds like people cheering for the last kid crossing the finish line. Sometimes it's ok to cheer for the last kid crossing the finish line because they're actually running a totally different race, and winning might look completely different. When I look at what Mistral does vs other organizations I'm impressed: They aren't profitable yet, but they're a lot closer than most and they're doing a hell of a lot with very little. Pointless racing story: I was in high school track with a really tough guy who was just not a runner. We went to a pretty messed up high school and if you screwed around in track practice sometimes the coach would make you run a crap race at the next meet, like steeplechase or hurdles. Well this guy and a few others screwed up and coach made them all run hurdles at a meet. He hooked every single one and fell on his face. Every time he got back up and kept on running. By the time he hit the finish line his knees were bleeding halfway down to his ankles. We cheered like hell and he was smiling ear to ear. Coach quit punishing us with races after that. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | badatnames 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
> This is a pretty grim prognosis for European AI. I think it's an incomplete read. What's the point in competing for a sizeable percentage of your funding when the finish line is incrementally being moved each month? Better spend it on leapfrogs which they seem to have done. Meanwhile Mistral have a natural ace in their pocket with respect to regulation in the form of CADA and the Cloud Sovereignty Framework. I can't think of another company that would qualify as SOV-3 under that regime | |||||||||||||||||||||||
| ▲ | simjnd 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
People were extremely dismissive of chinese models until recently. They went from 1 year behind frontier to 6 months behind frontier to 3 months behind frontier extremely fast. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | layer8 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
People are cheering for a kid that is gaining ground in an ongoing race. | |||||||||||||||||||||||
| ▲ | arrowleaf 4 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
Curious, why do you say the model is mediocre? I haven't tried it, so I can't pass any judgement... I've learned to distrust benchmark rankings. Are benchmarks and Artificial Analysis the yardstick you're using? | |||||||||||||||||||||||
| ▲ | epolanski 26 minutes ago | parent | prev | next [-] | ||||||||||||||||||||||
Doesn't matter, it's excellent, and it's European with European inference, which solves the pains of all my clients trying to build data lakes and processes on top of it. Nobody in the real world cares about minor benchmark differences in money losing coding agents. And nobody in the real world is giving Altman or Musk their data. | |||||||||||||||||||||||
| ▲ | danny_codes 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
Neat. Wait 3 months for the landscape to change entirely. LLM development is jumpy. It’s hard to extrapolate very far ahead. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | verdverm 4 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
Reminds me of Gemini 3.5 Pro | |||||||||||||||||||||||
| ▲ | porridgeraisin 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
First few models will always be slow improving and worse. The way to improvement is working your way through a gajillion evals [1], finding bugs, gaps, and curating training data (this part involves human design as well as raw inference compute) to fix it. This is very time intensive and can't easily be "done once and then everyone has lesser work to do" since every model is different. Well, one way to accelerate it is to simply have more compute, which mostly openai and anthropic have[2]. This is mistrals first 1T-scale model and I expect the 4th or 5th generation to be close to the best for many purposes. [1] These evals differ from the public ones like terminal-bench, are sometimes model-specific, need real, diverse usage to actually create, and are held secretly since quality of eval is the first driver behind the next step improvement of a model. [2] It is not close. This model was trained on less than 4k GPUs, whereas astra used north of 100k GPUs. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | Marciplan 2 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
hurr durr | |||||||||||||||||||||||
| ▲ | ismailmaj 4 hours ago | parent | prev [-] | ||||||||||||||||||||||
Those are comments from Europe. The US is waking up now and I expect them to be much harsher. I really want them to win as that's our last horse in the AI race, but ~200 research-oriented devs out of 1800 employees? I believe they agree it's pretty doomed and have pivoted. | |||||||||||||||||||||||
| |||||||||||||||||||||||