← Back to context

Comment by andai

9 hours ago

Wait, what does that number mean? I thought it always uses the cache price when the prefix matches.

When the prefix matches a request sent to the same Providor. The thing is the TTL is different for each provider, some cache for 5 minutes some cache for 1hr. Its ideal to only use one provider per agent session / and per model with the best cache hit % if you care about costs.