0
I’m considering prompt caching for a Claude API workflow with a large, mostly unchanged system prompt. Does caching mainly affect the cost of repeated requests, or can it also change how the model handles the cached context?
I’d like to hear what developers check around cache duration, prompt changes, and whether the feature is worthwhile when requests share only part of their context.