Comment by mswphd
2 hours ago
I've heard that certain inference providers may have different quality of caching implementations, so even if the listed numbers are as you say, the practical cache hit % you get might be significantly different/incur significantly different costs.
No comments yet
Contribute on Hacker News ↗