Remix.run Logo
neevans a day ago

tbh even if its better model due to lot of restrictions its not that useful than opus.

NitpickLawyer a day ago | parent | next [-]

100% this. There's currently this [1] submission that hasn't gained much attention, but is really important. In this [2] incident report from HuggingFace, they talk about detecting an attack and not being able to analyse the logs / IoC with API models because of guardrails. If not even highly regarded reputable companies can't sort out access to SotA models for blue team use, the raw capabilities don't matter. They're useless paperweights (hah!), and nothing else. Having to resort to open models is insane!

[1] - https://news.ycombinator.com/item?id=48965243

[2] - https://huggingface.co/blog/security-incident-july-2026

throwa356262 a day ago | parent [-]

Key part from [2]:

"When we started the log analysis, we first used frontier models behind commercial APIs. This did not work [...] We ran the forensic analysis instead on GLM 5.2, an open-weight model, on our own infrastructure. [...] The practical lesson for defenders: have a capable model you can run on your own infrastructure vetted and ready before an incident, both to avoid guardrail lockout [...]"

hodgehog11 a day ago | parent | prev [-]

Tell that to my colleagues. Despite Sol getting the attention, Fable is really starting to have an impact on mathematicians right now. It has unbelievable insights in a lot of cases that can rapidly speed up progress.