We're Probably Caching the Wrong Things in AI
Most AI caching today focuses on things like:
Embeddings
Retrieved documents
Prompt templates
Final responses
All of these are valuable.
But they have one thing in common.
They optimize around t
coalent.hashnode.dev1 min read