Question 263
A senior developer is optimizing their token costs. They discover that 60% of their API cost comes from input tokens (most content is repeated context). Only 15% comes from output tokens, and 25% from thinking tokens. What optimization provides the biggest cost reduction?