| ▲ | nolok an hour ago | |||||||
I work in the same field for one of my company, in europe, and if you're not self hosting sorry but your worries are not something I can accept because models are very much not reliable on that front, let alone when you let the host decide HOW to serve a model (ressources allocated, different version of the same model, etc ...). I'm not being a d**, just saying, the problem you have is something that I have faced EXACTLY, and at least here it's not working until you host in house or remote but on raw hardware. Otherwise it keeps having subtle changes, and you will notice no LLM API providers has guarantees about these. | ||||||||
| ▲ | frde_me 40 minutes ago | parent [-] | |||||||
It's not that you're a d*, it's just that you lack any kind of nuance There's a whole spectrum between self-hosting open weight models and having a cloud provider swap models from under you Should you self host a model if want to maximize predictability to the limit? Yes. Does that mean it's wrong for someone hitting a model on API to expect that it won't switch to a completely different model under the hood from one day to another? Probably not. | ||||||||
| ||||||||