Cutting LLM inference costs by 36% with prompt caching
By lizakatz · 2026-09-02 · 1 points · 0 comments
https://www.neradot.com/post/cutting-inference-cost-36-percent-with-prompt-caching
By lizakatz · 1 points · 0 comments · on Hacker News, read on BetterNews.
Open the full discussion on BetterNews