| ▲ | simonw 5 hours ago | |
Depends on the quality of the results. These things are driven by text prompts. If it turns out the OpenAI one returns better quality results than open weight variants they'll be rewarded by the market. Anyone using a decision model like this is going to have to spin up their own evals - these are far harder to vibe-check than regular text output LLMs. | ||