| ▲ | IanCal an hour ago | |
How do caches work across models? I would have thought that was very model specific - if not I’ve really misunderstood what’s getting cached. | ||
| ▲ | armanckeser 35 minutes ago | parent [-] | |
I am not sure the author of the comment you are replying to understands that LLM systems have prompt caches | ||