Cutting LLM inference costs by 36% with prompt caching
Posted 2 hours ago by
lizakatz
1
points
https://www.neradot.com/post/cutting-inference-cost-36-percent-with-prompt-caching
0
comments