| ▲ | b112 an hour ago | |
Awesome, I used Claude to write a small python script to do the same with Linode's API. The only difference is I setup a persistent drive, and with Linode you can boot off of it. So my biggest start up lag is ~ 2 minutes to deploy + boot, then maybe 2 more to warm the model. I actually dislike LLMs. But I'm a realist, and on-demand compute like this is massive cost saving measure. (persistent drives are relatively cheap, compared to a box with several GPUs.. or even one. I find it worth the expense) | ||