| ▲ | TofuLover 6 hours ago | |||||||
More seriously though, I think we should be fine: we don't host any content, and what people do with the models is their own responsibility (legally speaking, in our jurisdiction, at least according to Claude -- we're talking to a real lawyer next week). Like any other provider, we offer no guarantees of sane, safe, or accurate results. | ||||||||
| ▲ | gguingff 5 hours ago | parent | next [-] | |||||||
Thanks for bringing up a service like this, it's quite important. A few serious questions if you don't mind. Confidentiality? Do you use any sort of logging and if not do you have a way to guarantee that your hosting providers are not snooping? Price vs Vast or Runpod? If i have a very large or a very small workload do you have a competitive rate vs a gpu provider that offers private gpu access? Subscription vs Api costs? Do you only offer api rate or will you offer discounted tokens for subscription? Subscription friendly towards open source harnesses such as omp? Heretic ablation vs other methods? KL divergence scores? Do you post train the weights yourselves or do you offer weights trained by other organizations and is this information available on the service? Cache hit/miss pricing policy? 90/10 or a different cache pricing policy, and how long do conversions stay in kv cache? Quantized cache and model? Do you offer a choice if i want a quantized model for speed or a quantized cache? If not do you publish the information? SGlang vs vllm or other inference engine? Do you publish your engine stack details? Thank you kindly I find the competition in this space very lacking. | ||||||||
| ||||||||
| ▲ | Xunjin 6 hours ago | parent | prev | next [-] | |||||||
>legally speaking, in our jurisdiction, at least according to Claude -- we're talking to a real lawyer next week That's going to be fun lol | ||||||||
| ||||||||
| ▲ | tiahura 4 hours ago | parent | prev [-] | |||||||
Let us know what your insurance is like. | ||||||||