Cutting LLM inference costs by 36% with prompt caching

  • Posted 2 hours ago by lizakatz
  • 1 points
https://www.neradot.com/post/cutting-inference-cost-36-percent-with-prompt-caching

0 comments