Economics

The cost math of LLM workloads, and where repeated calls quietly add up. 29 articles.

Subscribe with RSS

How to cut LLM API costs with semantic caching

The cheapest token is the one you never spend twice. Here's the simple math behind semantic caching, and where the savings actually come from.

Economics5 min read

More articles

Page 1 of 2