Cutting LLM inference costs by 36% with prompt caching(neradot.com)2 points by lizakatz 7 days ago | 0 comments