| ▲ | daybox 2 hours ago | |
> Never spend more than you budgeted I assume that this is "$ spent on search + $ spent on LLM" < budget, but how do you handle the LLM spending more than you would expect on a request? Or is this handled by max_tokens and some form of pricing table? (and if so, how does caching play a role?) | ||
| ▲ | lajosdeme 2 hours ago | parent | next [-] | |
[dead] | ||
| ▲ | 2 hours ago | parent | prev [-] | |
| [deleted] | ||