gptme will make much better use of Anthropic's prompt caching in the next release, significantly bringing down costs and latency! 馃コ
Been wondering why my Anthropic API costs are so high despite prompt caching, only to realize I wasn't using it right...
Serious case of skill issues, but they need to just make it automatic.



